cd /news/artificial-intelligence/openai-prepares-to-expand-ultrafast-… · home › topics › artificial-intelligence › article
[ARTICLE · art-140105] src=testingcatalog.com ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

OpenAI prepares to expand Ultrafast API to more users

OpenAI is preparing a wider rollout of its Ultrafast API mode around its DevDay event on September 29, with new references appearing across the OpenAI Platform and API documentation, according to TestingCatalog. The mode, already previewed with GPT-5.6 Sol at up to 750 output tokens per second and up to 14× faster inference than Standard, remains limited to selected customers and is powered by Cerebras. A hidden speed selector in the Responses API Playground would let developers choose between Standard, Fast, and Ultrafast processing, though OpenAI has not confirmed which GPT-6 models will support Ultrafast at launch.

by read1 min views1 publishedSep 26, 2026
OpenAI prepares to expand Ultrafast API to more users
Image: Testingcatalog (auto-discovered)

OpenAI appears to be preparing a wider rollout of its Ultrafast API mode around DevDay on September 29, with new references showing up across the OpenAI Platform and API documentation.

TestingCatalog spotted a dedicated speed selector being prepared for the Responses API Playground (currently hidden), where developers could choose between Standard, Fast, and Ultrafast processing. OpenAI has already officially previewed Ultrafast with GPT-5.6 Sol, describing speeds of up to 750 output tokens per second and up to 14× faster inference than Standard. Access remains limited to selected customers, and OpenAI confirmed that the mode is powered by Cerebras.

The timing makes broader availability during DevDay plausible. OpenAI has since released the GPT-6 family, including GPT-6 Sol and GPT-6 Astra, making support for these newer models a key thing to watch. No one has confirmed that every GPT-6 model will support Ultrafast at launch.

For developers, the main trade-off will likely be economics. Standard, Fast, and Ultrafast could let organizations choose latency based on each workload's value, reserving higher-cost inference for applications where response time directly affects revenue or productivity. The feature also fits OpenAI’s broader infrastructure strategy. The company announced a 750 MW Cerebras partnership earlier this year, with capacity being deployed in stages through 2028. TestingCatalog has separately spotted new OpenAI Platform onboarding tiers, including an Accelerate option aimed at production workloads. Together, these changes suggest DevDay could focus heavily on API infrastructure, compute tiers, and tools for developers building production AI systems.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @openai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/openai-prepares-to-e…] indexed:0 read:1min 2026-09-26 · —