Benchmarks

Citation-accuracy and reliability benchmarks for AI legal research and drafting tools, framed as risk assessment rather than product endorsement. Each entry discloses the benchmark source type (peer-reviewed academic study, vendor-published report, or independent test), the tool name and version tested, the test date, and a methodology summary, and visually separates independent findings from vendor-coordinated ones. Entries cross-link to any Risk Digest case naming the same tool. Does not contain case outcomes (Risk Digest) or how-to guidance (Workflows), though it links to both.

ToolVendorSourceHallucination rateVersion testedTest date
Andrew Truelove ex-con criminal history and legal timeline
Legal implications of Arizona's 2026 primary election results
Ashley Webb's Senate eligibility under Article I, Section 3
The D.C. Circuit's Biden Biographer FOIA Ruling Explained
California DMV test cheating legal consequences explained
Candidate Qualification Disputes Hinge on State Election Codes
Should Lawyers Use ChatGPT or Specialized Legal AI?
Cher v. Bono appeal tests the line between royalties and copyright
Chicago's Legal Layoffs Risk Linked to Budget Shortfall
Coca-Cola's Fairlife Breach Tests the SEC Materiality Clock
What Cornyn's PEPFAR Hold Reveals About the Appointments Clause
Lettuce Brands Safe from Cyclospora Still Face Legal Liability Risks
Why the D4vd Preliminary Hearing Matters for Probable Cause Practice
A Mississippi noise complaint is testing AI infrastructure's legal limits
Flock Safety Civil Liberties Abuses Create Legal Risk for Cities
How the Heppner Ruling Changes Privilege Protection for AI Prompt Logs
ICC Arrest Warrants for Israeli Leaders: The Compliance Crisis
Assessing Iran's execution surge under international criminal law
Who Owns the Jacobian Disproof That Claude AI Found?
Jason Alexander Apology and the Legal Consequences of Minor Marriage
Blogarama - Blog Directory