AI Labs Sell Legal Tools as Courts Keep Punishing Fake Citations
Google, Anthropic and SpaceXAI are selling AI to law firms even as courts fine lawyers for fake citations — nearly 2,000 such cases are now cataloged worldwide.

Updated
Why it matters
- Three of the four major AI labs launched or pitched legal products this year; Google's Gemini Enterprise for Legal entered preview in late August with Gottlieb, Freshfields, Weil and Williams & Connolly.
- Ontario lawyer Shahryar Mazaheri was ordered to pay $31,150 in June for Grok-drafted filings with fabricated citations; researcher Damien Charlotin has cataloged nearly 2,000 such cases worldwide.
- Stanford RegLab's 2024 study found hallucination rates of 17 and 33 percent on purpose-built legal AI platforms from LexisNexis and Westlaw; no independent audit covers Google's or Anthropic's new tools.
Three of the four major AI labs have launched or pitched products aimed at lawyers this year, even as courts continue to fine attorneys for filing AI-drafted documents packed with fabricated case law.
Google put Gemini Enterprise for Legal into preview in late August with four of the world's biggest firms on board. Anthropic released what is now known as Claude Legal Solutions three months earlier. Elon Musk wants in too: his SpaceXAI lab has put up a dedicated legal solutions page on its website.
The timing is awkward, to put it mildly. The same month Google announced its preview, the Ontario Law Society Tribunal ordered lawyer Shahryar Mazaheri to pay $31,150 for filing documents drafted with Grok that contained fake case citations. Researcher Damien Charlotin's database has cataloged nearly 2,000 such cases worldwide, involving more than 800 lawyers and 1,100 self-represented litigants.
The sanctions keep coming
The Ontario Law Society found last December that Mazaheri had used an earlier version of Grok to draft a factum — a written legal argument — for a client's appeal. Tribunal documents say the output included fabricated citations and invented legal principles. In the June costs decision, adjudicators ordered Mazaheri to pay $31,150 and wrote that an LLM "does not appreciate nuance or exercise judgment or use a moral compass," and is "strongly predisposed to giving an answer — any answer — rather than admitting ignorance." They described his AI-generated materials as "gibberish."
Elite firms are not immune. Sullivan & Cromwell, one of the highest-grossing law firms in the United States, pays its first-year attorneys a $225,000 base salary — and still filed an emergency letter with the Southern District of New York in April after a bankruptcy motion it submitted to Chief Judge Martin Glenn was found to contain 42 AI-generated fabrications.
The trend line is steep. In Canadian courts alone, reported cases involving fabricated citations rose from seven in 2024 to 86 in 2025, with 39 more in the first quarter of 2026.
This is not a new problem. Before any major AI company had released a dedicated legal tool, DoNotPay dropped plans in early 2023 to have an AI "robot lawyer" argue a case in court after state bar associations threatened criminal prosecution.
Built around verification
Google and Anthropic both position their legal products as designed for verification rather than free generation. Google's product routes queries through connectors to external legal databases such as Everlaw and NetDocuments, instead of relying on the model's training data in isolation. Google Cloud CEO Thomas Kurian said at launch that "ensuring every aspect of these agentic workflows is accurate, factual, and grounded in legal authority is of critical importance." Gottlieb, Freshfields, Weil and Williams & Connolly — four of the world's largest law firms — are participating in the preview program.
Anthropic has taken a similar approach with Claude Legal Solutions. The product ships with a reported 20 connectors to legal platforms and 12 pre-built plugins. The lab claims a 90.9 percent score on the BigLaw Bench legal reasoning benchmark for Opus 4.7, and counts Freshfields and Quinn Emanuel among its customers.
The key difference between these tools and the Mazaheri case is architectural. Rather than asking a model to generate citations from memory — the failure mode that has tripped up hundreds of lawyers — these products pull citations from external, verified sources.
Grok's thinner résumé
SpaceXAI's Grok-powered legal product has weaker credentials. A Cursor blog post stated that Grok 4.5 was suited for "finance, legal work, or anything else you do on a computer," but SpaceXAI's later announcement for the model focused on "coding, agentic tasks and knowledge work" and made no mention of legal applications. The legal solutions page also does not disclose an architecture or mechanism for tracing citations to primary sources.
SpaceX officially acquired Cursor for $60 billion on August 15, bringing the code editor's legal marketing under the Grok umbrella.
Verification remains the lawyer's job
Independent evidence on hallucination rates in legal AI is sobering. A 2024 study by Stanford's RegLab found hallucination rates of 17 percent and 33 percent after testing purpose-built legal AI platforms from LexisNexis and Westlaw. No comparable independent audit of Google's or Anthropic's legal tools exists yet.
Bar association rules place the duty to verify citations on the attorney who files them, regardless of whether AI was involved. Mazaheri's $31,150 penalty was not imposed because he used AI; it was imposed because he failed to check the output. That obligation will not change, no matter how carefully a legal AI tool is designed or how many vendor safeguards ship with it.
For the labs, legal services represent a lucrative, high-stakes market where accuracy failures carry judicial consequences rather than mere reputational ones. Until independent audits test the new connector-based tools, every filing they help produce will rest on the same safeguard that failed nearly 2,000 times already: a lawyer reading the citations before the judge does.
Original: damiencharlotin.com
More from James Calloway
Show full bio
News editor covering industry trends and analytics at AI In Context.
152 articles
Related articles
- OpenAI Launches Astra for Law With 230 Million-URL Legal Search Index
- Nonprofit Sues OpenAI Over AI Agents' Hacking of Hugging Face
- Court Docs: AI Execs Knew Chatbots Threatened Journalism
- OpenAI Bans ChatGPT Accounts Behind Fake FBI Recovery Scam
- Microsoft's Own Scientist Called AI Scraping 'Largest Theft of Labor in History'