{"slug": "nvidia-says-vera-rubin-is-in-full-production-with-openai-set-to-deploy-at-scale", "title": "Nvidia says Vera Rubin is in full production, with OpenAI set to deploy at scale in Q3", "summary": "Nvidia confirmed Monday that its Vera Rubin platform is in full production, with OpenAI planning to deploy the system at scale in the third quarter. Ian Buck, Nvidia's vice president of accelerated computing, told reporters at Nvidia headquarters that systems are shipping to customers including OpenAI, CoreWeave, Google Cloud, Microsoft Azure, Meta, and Dell. CoreWeave reported that its NVL72 racks deliver ten times the token output of the previous generation, while Nvidia claimed the Vera CPU is nearly twice as fast as AMD's Turin chip on Python workloads.", "body_md": "#### TL;DR\n\n*Nvidia confirmed Vera Rubin is in full production, with OpenAI deploying at scale in Q3 and CoreWeave reporting ten times the token output.*\n\nIan Buck told reporters at Nvidia headquarters that CoreWeave is already measuring a tenfold increase in token output from the new NVL72 racks, while the Vera CPU outpaces AMD on Python workloads\n\n*Nvidia confirmed Vera Rubin is in full production, with OpenAI deploying at scale in Q3 and CoreWeave reporting ten times the token output.*\n\nNvidia confirmed on Monday that its Vera Rubin platform has reached full production, with Ian Buck, the company’s vice president of accelerated computing, telling reporters at Nvidia headquarters that systems are now shipping to customers including OpenAI, CoreWeave, Google Cloud, Microsoft Azure, Meta, and Dell. OpenAI plans to adopt Vera Rubin at scale during the third quarter, [according to Bloomberg](https://www.bloomberg.com/news/articles/2026-07-21/nvidia-touts-progress-getting-new-rubin-design-to-customers?srnd=phx-technology), which first reported the briefing.\n\nCoreWeave, one of the first cloud providers to receive the hardware, told Bloomberg that its NVL72 racks are delivering ten times the token output of the previous generation. The NVL72 is a full-rack system that pairs 72 Rubin GPUs with Vera CPUs and uses liquid cooling to eliminate internal cabling, a design change Nvidia demonstrated at the headquarters event. Buck said the cooling approach allows the company to remove copper connections that previously limited how tightly components could be packed together.\n\nThe Vera CPU is the piece Nvidia built to [replace the processors it previously bought from others](https://thenextweb.com/news/nvidia-vera-chip-anthropic-openai), and the company used the briefing to draw direct comparisons with AMD. Nvidia claimed the Vera CPU is nearly twice as fast as AMD’s Turin chip on Python workloads, a benchmark chosen because Python dominates the software stack that runs most AI inference. Anthropic and OpenAI are among the first labs to receive the processor, alongside Perplexity, SpaceX, and Oracle.\n\nThe production milestone comes at an awkward moment for Nvidia’s stock. The company’s shares have risen nine percent this year, while the broader chip index has climbed 66 percent over the same period, according to Bloomberg. Intel, ARM, and AMD have all more than doubled.\n\nAnalysts project Nvidia’s revenue will increase 82 percent to roughly $393 billion for the fiscal year, but that growth rate has not translated into the kind of share price momentum investors saw during the Blackwell cycle.\n\nNvidia has steadily built out the Vera Rubin story over the past two months. Jensen Huang [declared the platform in full production at Computex in early June](https://thenextweb.com/news/jensen-huang-computex-2026-keynote) and named Anthropic, OpenAI, SpaceX, and Oracle as early recipients. Monday’s briefing added the performance numbers and customer testimonials that the keynote stage did not provide, turning a claim into a set of measurable benchmarks.\n\nThe numbers Nvidia chose to highlight are its own and its customers’ rather than independent tests, a distinction worth noting. CoreWeave’s tenfold token figure and the Python benchmark against AMD came from the company’s own briefing, not from a third party. Volume shipments across all named customers, and independent verification of the performance claims, are still ahead.\n\nGet the most important tech news in your inbox each week.", "url": "https://wpnews.pro/news/nvidia-says-vera-rubin-is-in-full-production-with-openai-set-to-deploy-at-scale", "canonical_source": "https://thenextweb.com/news/nvidia-vera-rubin-full-production-customers", "published_at": "2026-07-21 16:03:10+00:00", "updated_at": "2026-07-21 16:49:48.987225+00:00", "lang": "en", "topics": ["artificial-intelligence", "ai-chips", "ai-infrastructure", "ai-products"], "entities": ["Nvidia", "OpenAI", "CoreWeave", "Google Cloud", "Microsoft Azure", "Meta", "Dell", "AMD"], "alternates": {"html": "https://wpnews.pro/news/nvidia-says-vera-rubin-is-in-full-production-with-openai-set-to-deploy-at-scale", "markdown": "https://wpnews.pro/news/nvidia-says-vera-rubin-is-in-full-production-with-openai-set-to-deploy-at-scale.md", "text": "https://wpnews.pro/news/nvidia-says-vera-rubin-is-in-full-production-with-openai-set-to-deploy-at-scale.txt", "jsonld": "https://wpnews.pro/news/nvidia-says-vera-rubin-is-in-full-production-with-openai-set-to-deploy-at-scale.jsonld"}}