Benchmarks

Citation-accuracy and reliability benchmarks for AI legal research and drafting tools, framed as risk assessment rather than product endorsement. Each entry discloses the benchmark source type (peer-reviewed academic study, vendor-published report, or independent test), the tool name and version tested, the test date, and a methodology summary, and visually separates independent findings from vendor-coordinated ones. Entries cross-link to any Risk Digest case naming the same tool. Does not contain case outcomes (Risk Digest) or how-to guidance (Workflows), though it links to both.

ToolVendorSourceHallucination rateVersion testedTest date
General-Purpose AI vs. Purpose-Built Contract Review: What the 2026 Benchmarks Mean for Your Professional Responsibility
What Virginia HB 1479 Means for Hit-and-Run Punitive Damages
How international law fails in the Strait of Hormuz tanker crisis
House rejects transgender military ban amendment for 2027 NDAA
Has the imminence doctrine justified the US airstrikes on Iran?
What Legal Framework Applies to Kalshi vs. Sportsbooks
Kalshi World Cup Sports Contracts Legal Status by State
New York Landlord Liability After Two Daycare Fentanyl Busts
Legal AI: Time Saved, Profits Unchanged?
No, Spain's Property Law Doesn't Fine You 3,000€ for a Flag
Limits and Liabilities: A Professional Responsibility Framework for AI Contract Review in 2026
The Legal Mechanics of a Long John Silver's Franchisee Bankruptcy
The Manson-MKUltra Link Fails Every Legal Standard
Mata v. Avianca: The ChatGPT Citation Hallucination That Led to Court Sanctions (2023)
When is a car not road-legal? The McLaren Senna case
MV Barima death toll reaches 41, legal inquiry underway
What New York's Data Center Moratorium Means for AI Development
The Asylum Paradox for North Korean Defectors
Law Firm AI Governance After the OpenAI–Hugging Face Breach
Peggy Flanagan and the legal path to Medicare for All
Blogarama - Blog Directory