# The Next Evolution of AI Infrastructure: Inside the Architecture Powering the AI Factory Era

> Source: <https://news.lenovo.com/inside-the-architecture-powering-the-ai-factory-era/>
> Published: 2026-07-23 18:30:25+00:00

AI factories aren’t the future anymore; they’re becoming the new blueprint for hyperscale AI. As enterprises and cloud providers move beyond deploying individual AI servers toward building AI factories capable of supporting massive inference and training workloads, infrastructure must evolve from standalone systems to integrated rack-scale designs.

That shift took center stage at Lenovo Tech World ’26, where AMD Chair and CEO Dr. Lisa Su announced Lenovo as one of the earliest OEM adopters of the AMD Helios™ rack-scale solution. Lenovo is applying decades of infrastructure engineering expertise to help customers build the next generation of AI factories – from scale-up and scale-out cluster design to the deployment and lifecycle services that keep AI environments running at scale.

Designed for the next era of generative and agentic AI, AMD Helios™ brings together next-generation AMD Instinct™ GPUs, AMD EPYC™ CPUs, AMD Pensando™ networking technologies and the AMD ROCm™ software ecosystem into an open, rack-scale architecture built on Open Compute Project (OCP) and Open Rack Wide (ORW) specifications. Engineered for hyperscale inference, training and fine-tuning, the solution is built to scale from rack-level systems into datacenter-scale clusters for distributed inference and the largest foundation-model training runs.

To support AI at this scale, the platform combines **72 AMD Instinct™ MI455X GPUs**, next-generation AMD EPYC™ processors, **31TB of HBM4 memory**, and AMD Pensando™ networking technologies in a single rack. The result is **up to 2.9 exaFLOPS of FP4 AI inference performance** and **1.4 exaFLOPS of FP8 AI training performance**, giving hyperscalers and NeoCloud providers the performance needed to train, fine-tune, and serve increasingly sophisticated AI models.

“AI infrastructure is rapidly evolving from standalone servers to fully integrated rack-scale clusters,” said Conor Malone, Vice President and General Manager of Cloud Service Providers, Lenovo. “Working with AMD, we’re bringing together world-class cluster design with Lenovo’s expertise in deploying at scale, helping customers leverage this solution to build AI environments that are open, scalable, and ready for what’s next.”

**More Than Hardware: A Faster Path to Deployment**

Preparing organizations for the new era of AI isn’t simply about deploying new hardware. It requires expertise in cluster design, liquid cooling, workload optimization, and lifecycle management. Leveraging its proven experience in deploying large-scale AI infrastructure, Lenovo helps customers accelerate deployment, optimize performance, reduce time to first token, and move AI environments into production faster.

Lenovo’s engineering and services experts work alongside customers to assess workload requirements, right-size deployments, optimize GPU configuration, and tune performance for the workload at hand. From there, ongoing full-lifecycle services capabilities reduce stand-up time and support long-term differentiation with greater speed and efficiency, helping organizations maximize GPU utilization and reduce deployment risk.

The AI factory may be changing how infrastructure is built, but Lenovo’s goal remains the same: help organizations deploy AI where it delivers the greatest value. Through its Hybrid AI vision, Lenovo is building the open, scalable infrastructure that enables customers to run AI across edge, enterprise, cloud, and hyperscale environments—meeting them where they are today while helping them prepare for what’s next. Lenovo solutions based on the AMD Helios rack-scale solution are expected to be available in 4Q 2026.
