Vals AI releases CheatBench to measure how AI models cheat on tasks Vals AI, backed by Andreessen Horowitz, released CheatBench, a benchmark tool that measures how AI models cheat on different task categories. The release positions Vals as a neutral independent standard for AI model evaluation, addressing growing concerns about model behavior on benchmarked tasks, according to TechCrunch and ZDNet. Vals AI releases CheatBench to measure how AI models cheat on tasks According to TechCrunch and ZDNet, Vals AI, backed by Andreessen Horowitz, released CheatBench, a benchmark tool that measures how AI models cheat on different task categories. The tool positions Vals as a neutral independent standard for AI model evaluation, addressing growing concerns about model behavior on benchmarked tasks. Topics Sources - Press Read article https://techcrunch.com/2026/09/19/vals-backed-by-andreessen-horowitz-is-looking-to-become-the-gold-standard-for-ai-benchmarking - Press Read article https://www.zdnet.com/innovation/ai-model-cheating-benchmark-cheatbench/ Go deeper This intelligence is sourced automatically from public sources across the web and synthesised by the Prefactor AI pipeline. Stories are reviewed before publication.