OpenAI Will Limit Access to New Astra Model’s Cybersecurity Features OpenAI will limit access to the cybersecurity features of its new Astra AI model, releasing it 'soon' with advanced cyber capabilities restricted to a group of testers before expanding through its Daybreak Blue program for defensive purposes. The company said the model reaches its 'critical cybersecurity threshold,' capable of identifying and developing zero-day exploits without human intervention, and has added guardrails including monitoring and automatic stopping of unauthorized activity. This follows an August pause and an incident where OpenAI's AI models inadvertently hacked Hugging Face Inc.'s system in July, raising concerns about AI agents running amok. Bloomberg -- OpenAI plans to soon roll out a powerful new artificial intelligence model called Astra, but said it will limit who can use the software's most cutting-edge cybersecurity capabilities. Most Read from Bloomberg - Global Bond Selloff Sends Yields to the Highest Level Since 2008 https://www.bloomberg.com/news/articles/2026-09-01/australian-benchmark-bond-yield-jumps-to-level-last-seen-in-2011?utm campaign=bn&utm medium=distro&utm source=yahooUS - BYD Asked About Taking Over Canada Stellantis Plant, Mayor Says https://www.bloomberg.com/news/articles/2026-08-31/byd-asks-about-buying-stellantis-stla-brampton-plant-for-buses-mayor?utm campaign=bn&utm medium=distro&utm source=yahooUS - US-Iran Tensions High After Strikes on Two Tankers in Hormuz https://www.bloomberg.com/news/articles/2026-09-01/two-oil-supertankers-hit-by-projectiles-in-hormuz-marisks-says?utm campaign=bn&utm medium=distro&utm source=yahooUS - India Rejects Hague Ruling on Indus Waters Treaty With Pakistan https://www.bloomberg.com/news/articles/2026-09-01/india-ordered-to-uphold-pakistan-water-treaty-curb-kashmir-dam?utm campaign=bn&utm medium=distro&utm source=yahooUS In a blog post Tuesday, OpenAI said it will release the AI model "soon," without detailing exactly when, and that at first the model's ability to carry out advanced cybersecurity-related tasks will be limited to a group of testers. After that, the company will offer a larger pool of users access for defensive cybersecurity purposes through its Daybreak Blue program, which lets approved testers use its most capable models along with safeguards that are meant for use with cybersecurity work. The company had said in August that it was pausing some internal work on the Astra model in order to incorporate stricter safeguards, after it was found to be significantly capable at cybersecurity tasks. On Tuesday, OpenAI said it believes the model reaches its "critical cybersecurity threshold," meaning it's capable of identifying and developing zero-day exploits without human intervention. The company said it has increased the Astra model's guardrails to prevent it from being misused, particularly for cybersecurity-related actions. These safeguards include monitoring the model for unauthorized behavior during internal deployments and automatically stopping potentially unauthorized activity. In late August, the company released a report in which it said it could have reacted sooner to prevent an inadvertent hack that its AI models carried out on Hugging Face Inc. in July. That "unprecedented" incident, along with several other recent cybersecurity breaches, have ignited concerns about AI agents running amok. It has prompted some technology and government leaders to renew calls for curbs on the technology. OpenAI said in July that the models involved in the hack broke into Hugging Face's system, which hosts AI models and datasets, during an evaluation of their cyber capabilities. The models were operating without the usual safety guardrails at the time, the company has said, because OpenAI had intended them to remain in a testing area known as a "sandbox" — essentially, an isolated virtual software environment that's meant to run security tests or analyze unsafe code in a controlled way.