{"slug": "how-nvidia-gpus-help-accelerate-openais-gpt-6-astra-ultrafast", "title": "How NVIDIA GPUs Help Accelerate OpenAI’s GPT-6 Astra Ultrafast", "summary": "OpenAI released GPT-6 Astra Ultrafast in its API and to eligible ChatGPT Work and Codex users, delivering up to 8x faster token generation than Astra Standard mode on NVIDIA Blackwell GPUs, NVIDIA announced. OpenAI inference lead Philippe Tillet said NVIDIA's tooling investment let OpenAI make its models \"exceptionally good at programming Blackwell and Rubin GPUs,\" and OpenAI chief technology officer of compute Uday Ruddarraju said OpenAI used its internal models to optimize inference on NVIDIA GPUs. The faster generation is aimed at shortening coding agents' edit-test-debug cycles and reducing latency between tool calls.", "body_md": "GPT-6 Astra Ultrafast, running on [NVIDIA Blackwell GPUs](https://www.nvidia.com/en-us/data-center/technologies/blackwell-architecture/), is available now in the OpenAI API and to eligible ChatGPT Work and Codex users. \n\nAccelerated by inference optimizations through OpenAI’s models that tap into the capabilities of the NVIDIA Blackwell architecture, Ultrafast offers up to 8x faster token generation than the Astra Standard mode. For developers, faster generation can shorten coding agents’ edit-test-debug cycles, reduce the time spent generating responses between tool calls and make interactive applications feel more responsive.\n\nA faster response matters most when it’s repeated across a workflow: an agent writes code, uses a tool, checks the result and decides what to do next. Ultrafast brings Astra’s capabilities into these time-sensitive loops. NVIDIA AI infrastructure helps OpenAI serve more useful model outputs when developers need it.\n\n“NVIDIA’s deep investment in tooling and documentation has enabled us to make our models exceptionally good at programming Blackwell and Rubin GPUs,” said Philippe Tillet, inference lead at OpenAI. “Astra can turn that knowledge into high-performance kernels that make NVIDIA hardware compelling across the full frontier of latency, throughput and cost. With Astra Ultrafast, that means faster model responses as agents write code, use tools and work through complex tasks.”\n\n## **Continually Improving Performance**\n\nPerformance gains don’t stop when a model is deployed. OpenAI is using its own models to help refine the inference software running on NVIDIA GPUs, taking advantage of the platform’s programmability to test and implement improvements. That ongoing work can make model responses faster and deployed infrastructure more productive over time.\n\n“Our work with NVIDIA is helping us make AI faster and more useful,” said Uday Ruddarraju, chief technology officer of compute at OpenAI. “We used our internal models to optimize inference on NVIDIA GPUs, and NVIDIA’s programmability helped us deliver the acceleration behind Astra Ultrafast.”\n\nA programmable NVIDIA platform allows developers and researchers to reuse infrastructure across training, inference and reinforcement learning as models evolve. That flexibility helps teams repurpose compute resources as demand changes, improving utilization and avoiding overprovision for each workload.\n\n*Developers can use GPT-6 Astra Ultrafast through the API today. See the* *Ultrafast guide* *for access, pricing and implementation details.*", "url": "https://wpnews.pro/news/how-nvidia-gpus-help-accelerate-openais-gpt-6-astra-ultrafast", "canonical_source": "https://blogs.nvidia.com/blog/gpus-openai-gpt-6-astra-ultrafast/", "published_at": "2026-10-01 23:44:13+00:00", "updated_at": "2026-10-01 23:46:24.509479+00:00", "lang": "en", "topics": ["large-language-models", "ai-infrastructure", "ai-chips", "ai-agents", "ai-products"], "entities": ["NVIDIA", "OpenAI", "GPT-6 Astra Ultrafast", "NVIDIA Blackwell", "ChatGPT Work", "Codex", "Philippe Tillet", "Uday Ruddarraju"], "also_reported_by": [], "alternates": {"html": "https://wpnews.pro/news/how-nvidia-gpus-help-accelerate-openais-gpt-6-astra-ultrafast", "markdown": "https://wpnews.pro/news/how-nvidia-gpus-help-accelerate-openais-gpt-6-astra-ultrafast.md", "text": "https://wpnews.pro/news/how-nvidia-gpus-help-accelerate-openais-gpt-6-astra-ultrafast.txt", "jsonld": "https://wpnews.pro/news/how-nvidia-gpus-help-accelerate-openais-gpt-6-astra-ultrafast.jsonld"}}