Identify the page
Include the canonical case or model URL and the exact statement that needs review.
About the project
CriminalBench is an independent editorial project that organizes publicly documented AI model behaviour into sourced case files, transparent cumulative points and a deliberately satirical watchlist.
Editorial scope
The archive helps readers discover safety research, controlled evaluations and public developer disclosures. It is not a scientific benchmark, legal assessment, safety certification or allegation that a developer committed a crime.
Read the scoring methodology →Every published case links to public evidence, with primary sources preferred.
Controlled, simulated, adversarial and real-world settings are not treated as equivalent.
Editors approve structured evidence factors; a public formula derives points and classifications.
Material corrections should identify the affected URL, the disputed fact and a reliable public source.
Corrections and contact
Include the canonical case or model URL and the exact statement that needs review.
Link a primary or otherwise reliable public source supporting the correction.
State whether the issue concerns facts, evaluation setting, attribution or editorial scoring.