NVIDIA Vera Rubin NVL72 Rack at Hot Chips 2026
NVIDIA Corp. detailed its Rubin GPU and Vera Rubin NVL72 rack at Hot Chips 2026, targeting agentic AI workloads with a full-stack platform spanning seven chips and five racks. The company claims up to…
NVIDIA Corp. detailed its Rubin GPU and Vera Rubin NVL72 rack at Hot Chips 2026, targeting agentic AI workloads with a full-stack platform spanning seven chips and five racks. The company claims up to…
NVIDIA announced the production release of the NVIDIA Groq 3 LPX, a dedicated interactive inference accelerator for the Vera Rubin NVL72 platform, achieving 3,400 tokens per second on the Gemma 4 31B …
SpaceXAI plans to launch Nvidia Vera Rubin NVL72 systems into orbit in late 2027, with significant scale targeted for 2028, to power its Starmind AI satellite and expand computing infrastructure for G…
SpaceX and Nvidia will launch a space-optimized version of Nvidia's Vera Rubin NVL72 AI supercomputer into low Earth orbit aboard the Starmind AI satellite in Q4 2027, with broader deployment planned …
NVIDIA announced that its NVIDIA Groq 3 LPX rack-scale system is now in full production, delivering 3,400 output tokens per second on a Gemma 4 31B benchmark for 100,000-token long-context use cases, …
NVIDIA's Vera Rubin NVL72 systems deliver up to 30x higher throughput per megawatt than the GB300 NVL72 on agentic AI workloads, according to new measured performance data from NVIDIA using the SemiAn…
SemiAnalysis's AgentX benchmark shows NVIDIA's Vera Rubin NVL72 delivers up to 30x higher AI-factory throughput per megawatt than GB300 NVL72 for agentic-coding inference, while GB300 NVL72 extends it…
NVIDIA announced that its Groq 3 LPX inference accelerator is now in full production, delivering a record 3,400 output tokens per second in Artificial Analysis benchmarking running Gemma 4 31B with a …
Microsoft has received the first production Vera Rubin systems from Nvidia, CEO Satya Nadella confirmed on August 21, 2026, making Microsoft one of the earliest adopters of Nvidia's next-generation AI…
NVIDIA has moved its 800-VDC power architecture from concept to production, with the first MGX-compatible 800-VDC power rack scheduled for Q3 2026, delivering high-voltage DC to AI compute racks while…
SpaceX announced a partnership with Nvidia to design the compute payload for its Starmind AI1 satellite, the first in a planned constellation for in-orbit AI workloads, with each satellite carrying Nv…
SpaceX and Nvidia announced on August 4, 2026, that they are jointly developing the compute payload for the Starmind AI1 satellite, a spacecraft designed to run AI workloads in low Earth orbit using N…
AMD CEO Lisa Su announced the Helios rack-scale AI system and sixth-generation Epyc 9006 Series CPUs at the Advancing AI conference in San Francisco on July 23, directly challenging Nvidia's dominance…
AMD launched the Instinct MI400 Series data-center GPU family, Helios rack-scale solution, and CDNA 5 architecture at its Advancing AI 2026 keynote, with the flagship Instinct MI455X featuring 320 bil…
AMD CEO Lisa Su announced a partnership with chip startup Cerebras to pursue disaggregated inference, splitting AI workloads across different hardware types, as AMD claims its Helios server system del…
Nvidia's Vera Rubin NVL72 delivers 5.4x performance per megawatt and 5x performance per dollar over GB200 NVL72 on DeepSeek R1 inference, according to early engineering samples from CoreWeave. The sec…
Analyst estimates from BofA Global Research and Morgan Stanley project an NVIDIA Rubin Ultra rack costing nearly $21 million, with HBM4e memory alone accounting for about $1.5 million per rack. The fi…
CoreWeave has deployed NVIDIA Vera Rubin NVL72 at rack scale, becoming the first cloud provider to validate the system. The company integrated training, continuous inference, observability, and autono…
Delta Electronics (Thailand) unveiled a prefabricated AI modular data center solution at COMPUTEX 2026 on June 5, 2026, integrating liquid cooling with an 800VDC power architecture to cut deployment t…
CoreWeave has developed two new innovations for AI data centers, a programmable valve system called 'Valvey' and a unified rack controller called 'Racky', designed to improve cooling and control for N…