cd /news/artificial-intelligence/amd-cerebras-partner-on-joint-helios… · home topics artificial-intelligence article
[ARTICLE · art-70607] src=sdxcentral.com ↗ pub= topic=artificial-intelligence verified=true sentiment=↑ positive

AMD, Cerebras partner on joint Helios rack-scale AI inference platform

AMD and Cerebras Systems announced a technical partnership to develop a joint Helios rack-scale AI inference platform, integrating Cerebras' Wafer-Scale Engine with AMD's Helios system to deliver up to five-times higher tokens per second per watt. The solution, available through Cerebras Cloud in the second half of 2026, targets ultra-low-latency inference workloads. Cerebras CEO Andrew Feldman said the partnership will bring ultra-fast inference to more customers.

read2 min views1 publishedJul 23, 2026
AMD, Cerebras partner on joint Helios rack-scale AI inference platform
Image: Sdxcentral (auto-discovered)

SAN FRANCISCO – AMD continued its Helios customer wins with a doozy of a partnership: wafer-scale chip maker Cerebras Systems.

Having already secured a tie-up with Anthropic at the start of the week, the chip giant revealed at this week's Advancing AI event that a technical partnership had been struck with the company that makes giant chips. The deal will see Cerebras deploy AMD Helios rack-scale systems in its data centers.

The pair also plan on developing a joint solution that would bring Cerebras’ Wafer-Scale Engine integrated as part of Helios, a combined platform they claim would provide up to five-times higher tokens per second per watt (T/s/W).

Set to become available initially through Cerebras Cloud in the second half of 2026, the joint offering would see Helios act as the high-throughput rack-scale system needed to process sizable complex requests, while Cerebras’s mammoth SRAM-based hardware would bring the ultra-low-latency performance needed to return tokens in what is essentially real time.

The resulting solution was described as being “designed specifically for the ultra-low-latency segment of the inference market, with AMD Helios as the foundation for high-throughput and balanced inference workloads across the data center.”

“The demand for ultra-fast inference is growing at an unprecedented pace. Cerebras delivers the world’s fastest, ultra-low-latency inference,” said Andrew Feldman, CEO and co-founder, Cerebras. “Partnering with AMD gives us an incredible opportunity to bring that performance to even more customers.”

Cerebras would join an ever-growing list of firms signed on to make use of AMD’s Nvidia NVL72 rival Helios, which already counts Meta, OpenAI, Oracle, and Tata Consultancy Services (TCS).

Earlier this week, Microsoft put pen to paper on a Helios rack-scale deal, while AMD revealed a $5 billion investment in Anthropic, with the Claude developer in turn making use of up to two gigawatts of AMD Instinct MI450 Series graphic processing units (GPUs) via Helios.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @amd 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/amd-cerebras-partner…] indexed:0 read:2min 2026-07-23 ·