cd /news/artificial-intelligence/amds-helios-puts-72-gpus-and-31-tera… · home topics artificial-intelligence article
[ARTICLE · art-65684] src=thenextweb.com ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

AMD’s Helios puts 72 GPUs and 31 terabytes of HBM4 in one rack. It is AMD’s answer to Nvidia’s NVL72.

AMD announced Helios, a rack-scale AI system packing 72 MI455X GPUs, 31 terabytes of HBM4 memory, and 2.9 exaflops of inference compute, directly competing with Nvidia's NVL72. Built on open standards like UALink and Ultra Ethernet, Helios targets hyperscalers and sovereign compute customers seeking vendor flexibility. Engineering samples ship in H2 2026, with mass production in Q2 2027.

read2 min views1 publishedJul 20, 2026
AMD’s Helios puts 72 GPUs and 31 terabytes of HBM4 in one rack. It is AMD’s answer to Nvidia’s NVL72.
Image: Thenextweb (auto-discovered)

TL;DR

AMD Helios packs 72 MI455X GPUs, 31TB HBM4, and 2.9 exaflops of inference into one rack. Built on open standards. Engineering samples H2 2026, mass production Q2 2027.

The rackscale system delivers 2.9 exaflops of FP4 inference and 1.4 exaflops of FP8 training. It uses open standards. Engineering samples ship this year.

AMD Helios packs 72 MI455X GPUs, 31TB HBM4, and 2.9 exaflops of inference into one rack. Built on open standards. Engineering samples H2 2026, mass production Q2 2027.

AMD’s Helios is a single rack containing 72 Instinct MI455X GPUs, 31 terabytes of HBM4 memory, and 2.9 exaflops of FP4 inference compute. It is AMD’s first rack-scale AI system and its direct answer to Nvidia’s Vera Rubin NVL72. The system uses 18 compute trays, each holding four MI455X accelerators on the new CDNA 5 architecture and one sixth-generation EPYC “Venice” CPU. Engineering samples ship in the second half of 2026. Mass production begins Q2 2027.

The architecture bet is open standards. Helios uses UALink for scale-up interconnect between GPUs within the rack, Ultra Ethernet Consortium specifications for scale-out networking between racks, and the OCP Open Rack Wide form factor. Nvidia’s competing NVL72 uses proprietary NVLink. AMD is betting that data centre operators who do not want to be locked into a single vendor’s interconnect will pay for the flexibility. AMD Pensando AI NICs handle the networking with programmable hardware and UEC-ready RDMA.

The numbers are designed to compete on memory, not just compute. Each MI455X GPU carries HBM4 with 19.6 TB/s of bandwidth. The full rack delivers 260 TB/s of scale-up bandwidth and 43 TB/s of scale-out bandwidth. That memory capacity matters for frontier model training and long-context inference, where the bottleneck has shifted from raw compute to how much data the system can hold and move. The AI-driven memory crisis has pushed HBM prices up sharply, and 31TB of HBM4 in a single rack represents an enormous materials cost that only hyperscaler and sovereign compute budgets can absorb.

Supermicro showed Helios hardware at Computex in June. AMD committed billions to UK AI infrastructure at London Tech Week, and Helios is the hardware those commitments will run on. The ROCm software stack supports PyTorch, TensorFlow, and JAX, which means developers do not need to rewrite code to move from Nvidia’s CUDA ecosystem, at least in theory. Whether AMD can close the software gap that has kept it behind Nvidia in AI compute is the question Helios is designed to force. The hardware specs are competitive. The ecosystem is the test.

Get the most important tech news in your inbox each week.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @amd 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/amds-helios-puts-72-…] indexed:0 read:2min 2026-07-20 ·