cd /news/artificial-intelligence/openai-will-limit-access-to-new-astr… · home topics artificial-intelligence article
[ARTICLE · art-118227] src=ca.finance.yahoo.com ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

OpenAI Will Limit Access to New Astra Model’s Cybersecurity Features

OpenAI will limit access to the cybersecurity features of its new Astra AI model, releasing it 'soon' with advanced cyber capabilities restricted to a group of testers before expanding through its Daybreak Blue program for defensive purposes. The company said the model reaches its 'critical cybersecurity threshold,' capable of identifying and developing zero-day exploits without human intervention, and has added guardrails including monitoring and automatic stopping of unauthorized activity. This follows an August pause and an incident where OpenAI's AI models inadvertently hacked Hugging Face Inc.'s system in July, raising concerns about AI agents running amok.

read2 min views1 publishedSep 1, 2026
OpenAI Will Limit Access to New Astra Model’s Cybersecurity Features
Image: Ca (auto-discovered)

(Bloomberg) -- OpenAI plans to soon roll out a powerful new artificial intelligence model called Astra, but said it will limit who can use the software's most cutting-edge cybersecurity capabilities.

Most Read from Bloomberg

Global Bond Selloff Sends Yields to the Highest Level Since 2008 - BYD Asked About Taking Over Canada Stellantis Plant, Mayor Says - US-Iran Tensions High After Strikes on Two Tankers in Hormuz - India Rejects Hague Ruling on Indus Waters Treaty With Pakistan

In a blog post Tuesday, OpenAI said it will release the AI model "soon," without detailing exactly when, and that at first the model's ability to carry out advanced cybersecurity-related tasks will be limited to a group of testers. After that, the company will offer a larger pool of users access for defensive cybersecurity purposes through its Daybreak Blue program, which lets approved testers use its most capable models along with safeguards that are meant for use with cybersecurity work.

The company had said in August that it was pausing some internal work on the Astra model in order to incorporate stricter safeguards, after it was found to be significantly capable at cybersecurity tasks.

On Tuesday, OpenAI said it believes the model reaches its "critical cybersecurity threshold," meaning it's capable of identifying and developing zero-day exploits without human intervention. The company said it has increased the Astra model's guardrails to prevent it from being misused, particularly for cybersecurity-related actions.

These safeguards include monitoring the model for unauthorized behavior during internal deployments and automatically stopping potentially unauthorized activity.

In late August, the company released a report in which it said it could have reacted sooner to prevent an inadvertent hack that its AI models carried out on Hugging Face Inc. in July. That "unprecedented" incident, along with several other recent cybersecurity breaches, have ignited concerns about AI agents running amok. It has prompted some technology and government leaders to renew calls for curbs on the technology.

OpenAI said in July that the models involved in the hack broke into Hugging Face's system, which hosts AI models and datasets, during an evaluation of their cyber capabilities. The models were operating without the usual safety guardrails at the time, the company has said, because OpenAI had intended them to remain in a testing area known as a "sandbox" — essentially, an isolated virtual software environment that's meant to run security tests or analyze unsafe code in a controlled way.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @openai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/openai-will-limit-ac…] indexed:0 read:2min 2026-09-01 ·