cd /news/artificial-intelligence/openai-previews-ultrafast-gpt-5-6-so… · home topics artificial-intelligence article
[ARTICLE · art-95661] src=9to5mac.com ↗ pub= topic=artificial-intelligence verified=true sentiment=↑ positive

OpenAI previews ‘Ultrafast’ GPT-5.6 Sol running up to 14 times faster

OpenAI is previewing a new Ultrafast service tier that runs its GPT-5.6 Sol model up to 14 times faster than standard processing, generating up to 750 output tokens per second via Cerebras hardware. The tier launches first through the OpenAI API and is aimed at latency-sensitive tasks such as voice, customer support, commerce, developer agents, financial research, and security response, with access limited to select customers while OpenAI evaluates capacity.

read1 min views3 publishedAug 13, 2026
OpenAI previews ‘Ultrafast’ GPT-5.6 Sol running up to 14 times faster
Image: 9To5Mac (auto-discovered)

OpenAI is previewing a new way to run its most capable GPT-5.6 model at dramatically higher speeds. The company says its new Ultrafast service tier can run GPT-5.6 Sol up to 14 times faster than standard processing.

Ultrafast mode launches first through the OpenAI API and is powered by Cerebras. It can generate up to 750 output tokens per second, potentially bringing frontier-level performance to workflows where latency matters as much as model intelligence.

OpenAI introduced GPT-5.6 Sol in June alongside the balanced Terra and speed-focused Luna models. The full family became broadly available in July, including through ChatGPT, Codex, and the API.

The company sees Ultrafast supporting live or near-production tasks including voice, customer support, commerce, developer agents, financial research, and security response.

OpenAI says its own developers have used it to analyze logs and traces during incidents, as well as compress research cycles that previously ran overnight into multiple iterations during the workday.

Access remains limited to a select group of customers while OpenAI evaluates how the added speed changes real-world products and expands capacity.

Businesses can join the Ultrafast waitlist by sharing their workload, latency requirements, expected usage, and other details.

Do more with your Apple products

[Apple AirTag 2 | Add Find My tracking to keys, bags, bikes, more](https://amzn.to/4vjNWN1)

[AirPods 4 | Apple’s newest wireless headphones](https://amzn.to/43ndewV)

[AirPods Pro 3 | Apple’s best wireless headphones](https://amzn.to/4ut3FYp)

[Beats USB-A to USB-C Cable | The official CarPlay cable](https://amzn.to/4mpKs7K)

*FTC: We use income earning auto affiliate links.* [More.](https://9to5mac.com/about/#affiliate)

[our homepage](http://9to5mac.com/)for all the latest news, and follow 9to5Mac on

[exclusive stories](https://9to5mac.com/feature/exclusive/),

[reviews](https://9to5mac.com/guides/review/),

[how-tos](https://9to5mac.com/guides/how-to/), and

[subscribe to our YouTube channel](https://www.youtube.com/9to5mac)
── more in #artificial-intelligence 4 stories · sorted by recency
── more on @openai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/openai-previews-ultr…] indexed:0 read:1min 2026-08-13 ·