Sullivan & Cromwell submitted a hallucinated citation to a bankruptcy court. Weeks later, Big Law’s largest firms announced they’re going deeper.
“In litigation, an authoritative-sounding hallucination is worse than no answer.” — Jay Madheswaran, CEO of Eve
“The work product is far beyond what I would’ve done on my own — probably ever.” — Christopher Kercher, Quinn Emanuel, on building a litigation platform on Claude with no coding background
Anthropic released 20+ legal integrations creating what Legal IT Insider called an “orchestration layer for legal work”: a single AI interface that accesses Westlaw, iManage, DocuSign, Box, and specialist legal AI products simultaneously. A lawyer can now ask Claude to review a contract, pull authority from Westlaw, compare it against internal precedent, identify litigation risk, draft amendments, and route the document for signature. That is a multi-system agentic workflow with no independent verification at any step. Freshfields deployed Claude to thousands of users and is co-developing AI-native workflows with Anthropic. Thomson Reuters is simultaneously a Claude data connector and a seller of competing AI products. Legal is now the top power-user job function on Anthropic’s Cowork platform. Sullivan & Cromwell, a white-shoe firm with massive internal resources, was caught submitting a hallucinated citation to a bankruptcy judge just weeks before this announcement.
Anthropic legal webinar
Cowork job function
filing hallucinated citation
Grounding is a technical control. It is not independent verification. Anthropic’s connector architecture may reduce hallucinations by restricting sources, but the claim that it works is made by the vendor selling the product. Eve evaluates Claude against “24+ legal-specific scorers,” but Eve is built on Claude. When a judge sanctions a firm for a hallucinated citation, the question is not “did the vendor say their grounding works?” The question is “did anyone independent verify it?” AVAAS provides that independent verification. A model whose grounding architecture fails to prevent fabricated citations does not pass certification, regardless of what the vendor’s internal benchmarks report.
Lichtenberg, N. (May 12, 2026). Fortune. fortune.com · Hill, C. (May 13, 2026). Legal IT Insider. legaltechnology.com
This entry is one of 37 documented cases in the AVAAS evidence ledger, a public record of AI and automated-system failures with a verified source on every entry.
Every case here reached a person.
AVAAS certifies how AI systems behave at the decision point, with documented third-party evidence of conformity to a published standard.
Certify Your AI →