cd /news/artificial-intelligence/nvidia-says-vera-rubin-is-in-full-pr… · home topics artificial-intelligence article
[ARTICLE · art-67290] src=thenextweb.com ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

Nvidia says Vera Rubin is in full production, with OpenAI set to deploy at scale in Q3

Nvidia confirmed Monday that its Vera Rubin platform is in full production, with OpenAI planning to deploy the system at scale in the third quarter. Ian Buck, Nvidia's vice president of accelerated computing, told reporters at Nvidia headquarters that systems are shipping to customers including OpenAI, CoreWeave, Google Cloud, Microsoft Azure, Meta, and Dell. CoreWeave reported that its NVL72 racks deliver ten times the token output of the previous generation, while Nvidia claimed the Vera CPU is nearly twice as fast as AMD's Turin chip on Python workloads.

read3 min views1 publishedJul 21, 2026
Nvidia says Vera Rubin is in full production, with OpenAI set to deploy at scale in Q3
Image: Thenextweb (auto-discovered)

TL;DR

Nvidia confirmed Vera Rubin is in full production, with OpenAI deploying at scale in Q3 and CoreWeave reporting ten times the token output.

Ian Buck told reporters at Nvidia headquarters that CoreWeave is already measuring a tenfold increase in token output from the new NVL72 racks, while the Vera CPU outpaces AMD on Python workloads

Nvidia confirmed Vera Rubin is in full production, with OpenAI deploying at scale in Q3 and CoreWeave reporting ten times the token output.

Nvidia confirmed on Monday that its Vera Rubin platform has reached full production, with Ian Buck, the company’s vice president of accelerated computing, telling reporters at Nvidia headquarters that systems are now shipping to customers including OpenAI, CoreWeave, Google Cloud, Microsoft Azure, Meta, and Dell. OpenAI plans to adopt Vera Rubin at scale during the third quarter, according to Bloomberg, which first reported the briefing.

CoreWeave, one of the first cloud providers to receive the hardware, told Bloomberg that its NVL72 racks are delivering ten times the token output of the previous generation. The NVL72 is a full-rack system that pairs 72 Rubin GPUs with Vera CPUs and uses liquid cooling to eliminate internal cabling, a design change Nvidia demonstrated at the headquarters event. Buck said the cooling approach allows the company to remove copper connections that previously limited how tightly components could be packed together.

The Vera CPU is the piece Nvidia built to replace the processors it previously bought from others, and the company used the briefing to draw direct comparisons with AMD. Nvidia claimed the Vera CPU is nearly twice as fast as AMD’s Turin chip on Python workloads, a benchmark chosen because Python dominates the software stack that runs most AI inference. Anthropic and OpenAI are among the first labs to receive the processor, alongside Perplexity, SpaceX, and Oracle.

The production milestone comes at an awkward moment for Nvidia’s stock. The company’s shares have risen nine percent this year, while the broader chip index has climbed 66 percent over the same period, according to Bloomberg. Intel, ARM, and AMD have all more than doubled.

Analysts project Nvidia’s revenue will increase 82 percent to roughly $393 billion for the fiscal year, but that growth rate has not translated into the kind of share price momentum investors saw during the Blackwell cycle.

Nvidia has steadily built out the Vera Rubin story over the past two months. Jensen Huang declared the platform in full production at Computex in early June and named Anthropic, OpenAI, SpaceX, and Oracle as early recipients. Monday’s briefing added the performance numbers and customer testimonials that the keynote stage did not provide, turning a claim into a set of measurable benchmarks.

The numbers Nvidia chose to highlight are its own and its customers’ rather than independent tests, a distinction worth noting. CoreWeave’s tenfold token figure and the Python benchmark against AMD came from the company’s own briefing, not from a third party. Volume shipments across all named customers, and independent verification of the performance claims, are still ahead.

Get the most important tech news in your inbox each week.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @nvidia 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/nvidia-says-vera-rub…] indexed:0 read:3min 2026-07-21 ·