OpenAI slows the frontier to regain control OpenAI announced on Tuesday that it is slowing the development of its most powerful models, implementing safeguards across all stages of training after its systems breached Hugging Face during testing and its upcoming Astra model family reached the 'critical' threshold for cybersecurity capabilities. The company paused reinforcement learning training for two weeks and placed its largest planned frontier reinforcement learning run on hold, with CEO Sam Altman stating, 'Getting AI safety right is more important than any company’s momentum.' OpenAI is hitting the brakes on developing its most powerful and capable models. On Tuesday, the AI lab announced that it is implementing safeguards https://openai.com/index/pacing-model-development-cyber-capabilities/ "across all stages of the training process," which will slow the pace at which it's developing and scaling its models. The company said the move is necessary after its systems breached Hugging Face during testing, and after its upcoming family of models, Astra, reached the "critical" threshold for cybersecurity capabilities under its Preparedness Framework. Notably, in recent weeks, the company paused reinforcement learning training for two weeks while strengthening its research environments and expanding monitoring systems. Additionally, the largest planned frontier reinforcement learning run is "on hold" as it conducts smaller evaluations to test model behavior, the company said in its post. "Keeping increasingly capable systems aligned is a challenge the whole field will need to address," OpenAI said in its blog post. "The signals we are seeing from upcoming model progress make clear that we need a broader approach." The company will strengthen its safeguards in three key ways: - Monitoring, or detecting and responding to concerning behavior. For this, the company will expand "chain-of-thought monitoring" to escalate potential concerns. - Alignment, or reducing the likelihood of bad behavior before it happens. OpenAI said it intends to apply its core alignment techniques across more stages of the training process to better discourage unsafe behavior. - And security measures, or limiting what an AI system can access, such as raising security standards in the research environments where models are trained and evaluated. OpenAI CEO Sam Altman told reporter Alex Heath for TIME https://time.com/article/2026/08/18/openai-slowing-training/ that tugging on the reins has also redirected a significant amount of both compute and researchers to alignment research and monitoring systems. And although the Hugging Face incident made headlines, Altman also noted that the decision wasn't the result of just one incident, but rather its most powerful models showing "various degrees of misalignment" while advancing faster than expected. "Getting AI safety right is more important than any company’s momentum," Altman said. Our Deeper View The move to deliberately pause is a first for OpenAI, and largely unprecedented in an industry moving at lightning speed. Though Anthropic has talked about an industry-wide slowdown in a recent essay on recursive self-improvement https://www.thedeepview.com/articles/anthropic-s-rsi-warning-contrasts-with-ipo-filing , the company noted that a pause would have to be both industry-wide and international in order to keep one lab from overtaking the others. Additionally, slowing down is a popular move among the researchers and developers building AI, such as the coalition of more than 1,100 employees at OpenAI, Anthropic, Meta and Google https://www.thedeepview.com/articles/why-the-experts-building-ai-want-to-slow-it-down that signed a letter urging that the government and industry work together to "deliberately pace the frontier of automated AI development." It's commendable that OpenAI is listening to those pleas and keeping its models from causing more trouble. And of course, it's good PR. We'll see if OpenAI's safety slowdown influences other labs to follow suit.