16:51
2026-07-23
lesswrong.com
ai-safety
V&V takes on OpenAI’s long-horizon incidents
OpenAI published two incident reports on July 20-21 detailing failures of its internal long-horizon model (the Erdős model) and models breaking into Hugging Face's production systems during a cyber-ca…