cd /news/artificial-intelligence/openai-says-its-new-astra-ai-can-bui… · home topics artificial-intelligence article
[ARTICLE · art-118649] src=cryptonews.net ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

OpenAI says its new 'Astra' AI can build attacks without human help

OpenAI says its upcoming Astra model can find previously unknown software flaws and turn them into working attacks without human guidance, making it the first model classified as having 'Critical' cyber capabilities under its Preparedness Framework. In testing, Astra scored 100% on a benchmark for developing exploits from known vulnerabilities and found two previously unknown flaws, leading OpenAI to delay parts of its development and restrict its most advanced abilities to selected testers.

read2 min views1 publishedSep 2, 2026
OpenAI says its new 'Astra' AI can build attacks without human help
Image: Cryptonews (auto-discovered)

OpenAI says its upcoming Astra model can find previously unknown software flaws and turn them into working attacks without a human guiding each step, crossing a cybersecurity threshold that until recently belonged largely to expert hacking teams.

It is the first model OpenAI has classified as having “Critical” cyber capabilities under its Preparedness Framework, the firm wrote in a Tuesday post.

To qualify, a model must be able to find previously unknown software flaws, known as zero-days, and develop working exploits for them across hardened real-world systems without human intervention, or devise and execute an attack from little more than a high-level goal.

In testing, Astra scored 100% on a benchmark for developing exploits from known vulnerabilities and found two previously unknown flaws while building an exploit chain on a separate internal test.

It also broke out of a hardened browser sandbox and executed commands on the host computer, while separately finding and combining multiple flaws in an operating system to gain root access, OpenAI said.

The company has since delayed parts of Astra’s development while adding safeguards, and plans to initially restrict its most advanced cybersecurity abilities to selected testers.

That capability is particularly relevant to crypto, where a software flaw can be converted into money within minutes. CoinDesk reported in June that increasingly capable AI models could compress the work of searching code, finding misconfigurations and assembling attacks from days or weeks into machine-speed operations.

Read More: Crypto’s next billion-dollar hacker may move at superhuman speed

At the time, security researchers said the bigger change was not necessarily a new class of hack, but how quickly existing weaknesses could be found and exploited.

The advance follows other signs that frontier models are moving beyond answering questions and writing code. Anthropic’s Claude Fable 5 helped solve an 87-year-old mathematics problem in July, as CoinDesk reported.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @openai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/openai-says-its-new-…] indexed:0 read:2min 2026-09-02 ·