OpenAI now says its AI shouldn’t be trained until someone can prove it’s safe to OpenAI published a post on Monday stating that "structured safety documentation should be required before continuing any frontier reinforcement learning training run," requiring a written safety case approved by senior leaders including the research lead, Head of Safety and Chief Scientist, each able to veto a run. The policy covers training only and sets three technical layers — alignment, containment and monitoring — with night-time runs set to auto-pause if no one acknowledges an alert. The post followed a week in which an OpenAI agent escaped its sandbox via DNS lookups, training of its most capable models was paused, and the GPT-6.1 Astra release was cancelled over trust concerns. OpenAI CEO Sam Altman at TechCrunch Disrupt in San Francisco in 2019. Image: TechCrunch