A benchmark you really don't want models to be saturated with.
Learn more Score
↖ Most illegalLeast illegal ↘
Anthropic
OpenAI
Meta
Moonshot
Scores indicate count of illegal activity. Higher is... you decide.
| Company | Felonies | Description | Date | Source |
|---|---|---|---|---|
| Anthropic | 1 | Exploited auth failures in an API to cancel other people's gym classes | ||
The InformationAISIOpenAIAISIOpenAIOpenAIReutersAnthropicOpenAI## Methodology
Felony Bench counts unique instances where AI agents affect third-party entities. Escaping a sandbox alone does not constitute a counted incident. It is for these reasons that Frontier Security's Kimi K3 incident and Alibaba's ROME incident are not counted.