07:00
2026-10-06
metr.org
ai-safety
AI systems could cover up misbehavior
METR researchers found a vulnerability in about 10 minutes that could let an AI agent running inside an Inspect evaluation arbitrarily modify the transcript a human reviewer sees through the Inspect vā¦