AI on Trial: Legal Models Hallucinate in 1 out of 6 (or More) Benchmarking Queries
Artificial intelligence (AI) tools are rapidly transforming the practice of law. Nearly three quarters of lawyers plan on using generative AI for their work, from sifting through mountains of case law to drafting contracts to reviewing documents to writing legal memoranda. But are these tools reliable enough for real-world use? Large language models have a documented tendency to “hallucinate,” or make up false information. In one highly-publicized case, a New York lawyer faced sanctions for citing ChatGPT-invented fictional cases in a legal brief; many similar cases have since been reported. And our previous study of general-purpose chatbots found that they hallucinated between 58% and 82% of the time on legal queries, highlighting the risks of incorporating AI into legal practice. In his 2023 annual report on the judiciary, Chief Justice Roberts took note and warned lawyers of hallucinations. Across all areas of industry, retrieval-augmented generation (RAG) is seen and promoted as th
A new study reveals the need for benchmarking and public evaluations of AI tools in law. Artificial intelligence (AI) tools are rapidly transforming the practice of law. Nearly three quarters of lawyers plan on using generative AI for their work, from sifting through mountains of case law to drafting contracts to reviewing documents to writing legal memoranda. But are these tools reliable enough for real-world use? Large language models have a documented tendency to “ hallucinate ,” or make up false information. In one highly-publicized case, a New York lawyer faced sanctions for citing ChatGP
Explore this link on the map →related reading
- Hallucination (artificial intelligence) - Wikipediaen.wikipedia.org
- Training a State-of-the-Art Legal Agent with Harvey | Applied Computeappliedcompute.com
- Customizing models for legal professionals | OpenAIopenai.com
- Import AIjack-clark.net
- Why RAG won't solve generative AI's hallucination problem | TechCrunchtechcrunch.com
- What are AI hallucinations—and how do you prevent them?zapier.com
- Gabe Pereyra on X: "Model strategy for @harvey: We are working on the first model in our legal foundation model series, inspired by @cursor_ai's Composer. Two goals: 1. Allow us to serve frontier intelligence across our product surface areas at an affordable price and a strong security posture." / Xx.com
- Hallucination Mitigation using Agentic AI Natural Language-Based Frameworksarxiv.org
- Sam Altman says hallucinations are part of the "magic" of generative AI | IT Proitpro.com
- AI’s Hallucination Problem Isn’t Going Awayforbes.com
- AI #100: Meet the New Boss | Don't Worry About the Vasethezvi.wordpress.com
- The bitter lesson of LLM evalsparsed.com