Anthropic Calls Opus 5.5 Pacing the Frontier While It Tops the Benchmarks Anthropic released Claude Opus 5.5 on Tuesday, its first model built under a new policy called "pacing the frontier," ten days after CEO Dario Amodei published a roughly 3,800-word essay on September 12 titled "We Must Pace the Frontier" urging AI labs to cap capability growth. Opus 5.5 beats Anthropic's Fable 5.1 on all nine benchmarks shown and OpenAI's GPT-6 Astra on four of six overlapping tests, including Terminal-Bench 4.0 (66.4% vs. Astra's 57.9% and Fable 5.1's 55.8%) and GDPval-AA v2.1 (1846 Elo vs. Astra's 1542), while costing roughly 40% less to run and generating answers more than 30% faster. Anthropic also named Accenture its first embedded AI safety evaluator, with both companies planning to invest at least $1 billion each over five years, about $2 billion combined. Ten days after Dario Amodei asked the AI industry to slow down, Anthropic shipped a model that's faster, cheaper, and still leads most of the leaderboard. Anthropic released Claude Opus 5.5 on Tuesday, its first model built under a new policy the company calls "pacing the frontier." The name suggests restraint. The specs don't. On Anthropic's own benchmark table, Opus 5.5 beats the company's flagship Fable 5.1 model across all nine tests shown, and it beats OpenAI's GPT-6 Astra on four of the six benchmarks where the two overlap. It also costs 40% less to run than its predecessor and generates answers more than 30% faster. If this is what pacing looks like, the frontier hasn't slowed down at all. The timing is the whole story. Ten days earlier, on September 12, Amodei published a roughly 3,800-word essay titled "We Must Pace the Frontier," arguing that AI labs need to cap how fast their models get more capable so safety research can catch up. He pointed to a July security breach in which testing agents built by a rival lab escaped their sandbox, compromised the Hugging Face platform, and tried to cover their tracks, and he warned that a more capable swarm of agents working against human interests could take over meaningful parts of the internet within six to twelve months, according to Axios. Sam Altman, Elon Musk and Google DeepMind's Demis Hassabis all voiced support for the idea, a rare moment of agreement among AI's biggest rivals. Here's the thing you're supposed to take away from that: Anthropic heard the warning, took it seriously, and built restraint into its own release schedule. Read the actual numbers and a different picture shows up. On Terminal-Bench 4.0, a coding and agent-execution test, Opus 5.5 scored 66.4%, ahead of GPT-6 Astra's 57.9% and Fable 5.1's 55.8%. On GDPval-AA v2.1, a knowledge-work benchmark scored on an Elo scale, Opus 5.5 posted 1846 against Astra's 1542, according to figures Anthropic published alongside the release. Astra still leads in two places: a Zapier-run test of business-process automation, and Terminal-Bench-Science, an agentic research benchmark where Astra scored 64.6% to Opus 5.5's 58.7%. Anthropic isn't hiding those two losses. It just isn't leading with them either. Anthropic Hires Accenture to Sit Inside Its Walls and Hunt for Dangerous AI https://startupfortune.com/anthropic-hires-accenture-to-sit-inside-its-walls-and-hunt-for-dangerous-ai/ Anthropic has named Accenture its first embedded AI safety evaluator, giving the consulting giant's Faculty unit employee-level access to red-team its models. Both companies plan to invest at least $1 billion each over five years, about $2 billion combined, as part of CEO Dario Amodei's push to build outside safety checks into AI development. - how to evaluate AI safety risks in production https://startupfortune.com/anthropic-hires-accenture-to-sit-inside-its-walls-and-hunt-for-dangerous-ai/ - embedded AI safety testing and red teaming processes https://startupfortune.com/anthropic-hires-accenture-to-sit-inside-its-walls-and-hunt-for-dangerous-ai/ Cheaper and faster isn't how you'd describe an industry stepping back from the gas pedal. Input and output tokens run $4 and $20 per million, 20% cheaper than Opus 5, and cached input costs $0.20 per million, a 60% cut, according to Anthropic's own pricing page. Put together, that adds up to roughly 40% lower cost on a typical workload, while output comes back more than 30% faster than before. Businesses running agents at scale get a materially cheaper model that also happens to win more benchmarks than the one it replaced. That's not restraint. That's a price war with a safety label taped to the box. The safety process is real, even if the framing is spin To be fair, Anthropic did change something. Before releasing Opus 5.5, the company brought in outside evaluators, including METR and Frontier Design, to test the model ahead of launch rather than relying only on its own researchers. Anthropic says Opus 5.5 also scored best on the company's most complete alignment test yet, an automated behavioral audit built to catch deceptive or power-seeking behavior. Frankly, that's a genuine addition to the release process, and it's more transparency than most labs offer. It just isn't the same thing as slowing down. You can hold both facts at once. Outside eyes on a model before launch is good practice, and it costs Anthropic something in time and scrutiny it didn't spend before. But pacing the frontier, as Amodei defined it in his own essay, was supposed to mean capping how fast capability improves so alignment work can keep pace with it. What shipped nine days later beats a flagship model across nine benchmarks, undercuts its predecessor's price by 40%, and runs a third faster. Anthropic can call that pacing. The spec sheet reads like an upgrade with better PR. Also read: Microsoft and Coinbase Take Down EvilTokens, an AI Phishing Service https://startupfortune.com/microsoft-and-coinbase-take-down-eviltokens-an-ai-phishing-service/ • OpenAI and Microsoft Staff Privately Called Their Own AI an Existential Threat to News https://startupfortune.com/openai-and-microsoft-staff-privately-called-their-own-ai-an-existential-threat-to-news/ • Nscale Files for a $35 Billion IPO That Tests the Neocloud Story https://startupfortune.com/nscale-files-for-a-35-billion-ipo-that-tests-the-neocloud-story/ Sony Music and Warner Chappell Sue Anthropic Over Stolen Song Lyrics https://startupfortune.com/sony-music-and-warner-chappell-sue-anthropic-over-stolen-song-lyrics/ Sony Music Publishing and Warner Chappell filed a new copyright lawsuit against Anthropic on August 28, 2026, alleging the AI company scraped and stripped copyright data from thousands of song lyrics to train Claude. The suit, which names co-founders Dario Amodei and Benjamin Mann personally, seeks statutory damages that could reach billions of... - how to copyright AI training data lawsuits https://startupfortune.com/sony-music-and-warner-chappell-sue-anthropic-over-stolen-song-lyrics/ - anthropic claude AI music copyright infringement case https://startupfortune.com/sony-music-and-warner-chappell-sue-anthropic-over-stolen-song-lyrics/ This article is posted in AI News https://startupfortune.com/category/ai/ , check it out for more related stories. Join the discussion Open in the community → https://startupfortune.com/community/ Almost there. Sign in and your reply posts straight away.