Autonomous weapons explained in 101 seconds
For the UC Berkeley News “101 in 101” series, Stuart Russell, professor of electrical engineering and computer sciences and president of the International Association for Safe & Ethical… Video
At Black Hat USA 2026, OpenAI’s Eric Wallace (Alignment and Safety Research) and Michael Dalton (Security and Infrastructure) delivered a talk detailing how OpenAI had inadvertently…
Article
In his guest article in The Guardian, UC Berkeley professor Stuart Russell analyzes the significance of an open letter recently signed by more than 1,300 researchers and engineers from…
Report
The most dangerous outcomes of AI systems can be mitigated through red lines, governance instruments that prohibit unacceptable risks and behaviors. This Policy Paper describes the…
AI safety researchers at METR and Redwood Research present results from an independent investigation into the coordinated attack on Hugging Face by an agentic AI system developed by OpenAI….
Academic Paper
Human oversight is widely treated as a primary risk-mitigation measure, but the mere presence of a “human in the loop” is not sufficient for agentic AI systems, Margaret Mitchell, Samir…
Resource
Exhibit AI serves as an independent, public resource for monitoring legal cases involving AI companies and products. The platform provides comprehensive litigation guides, case trackers,…
While frontier AI models provide safeguards against misuse, there are vast differences between them, research by FAR.AI shows. The non-profit organization examined how hard it is to make… This press release by the International Association for Safe & Ethical AI (IASEAI) addresses a July 2026 security incident in which two agentic AI systems by OpenAI broke out of…
AI-generated inaccuracies and fabricated citations pose a risk to national security, Tech Policy Press contributor Rachel A. George argues. This is because government institutions across…