The Verge AI In July, OpenAI revealed that its AI agents had attacked Hugging Face without permission, sparking widespread concerns about AI safety. Since then, a string of similar incidents involving agents from Meta, Anthropic, Google, and other companies has fueled further fears about rogue AI. As disclosures implicating numerous AI models trickled out over the past few months, these seemed like separate incidents. But many share a common source: one specific company tasked with testing the agents.
Irregular, an Israeli startup that stress-tests AI models in "high-fidelity research platforms that simulate and monitor real-world AI security scenarios …
Get AI news in your inbox
Daily digest of what matters in AI.
Key Terms Explained #
AI Safety The broad field studying how to build AI systems that are safe, reliable, and beneficial.
Anthropic An AI safety company founded in 2021 by former OpenAI researchers, including Dario and Daniela Amodei.
Hugging Face The leading platform for sharing and collaborating on AI models, datasets, and applications.
OpenAI The AI company behind ChatGPT, GPT-4, DALL-E, and Whisper.