{"slug": "the-next-evolution-of-ai-infrastructure-inside-the-architecture-powering-the-ai", "title": "The Next Evolution of AI Infrastructure: Inside the Architecture Powering the AI Factory Era", "summary": "AMD Chair and CEO Dr. Lisa Su announced at Lenovo Tech World '26 that Lenovo is one of the earliest OEM adopters of the AMD Helios rack-scale solution, a platform combining 72 AMD Instinct MI455X GPUs, next-generation AMD EPYC processors, 31TB of HBM4 memory, and AMD Pensando networking in a single rack to deliver up to 2.9 exaFLOPS of FP4 AI inference performance and 1.4 exaFLOPS of FP8 AI training performance. Lenovo will apply its infrastructure engineering expertise to help customers build AI factories, with solutions based on AMD Helios expected to be available in 4Q 2026.", "body_md": "AI factories aren’t the future anymore; they’re becoming the new blueprint for hyperscale AI. As enterprises and cloud providers move beyond deploying individual AI servers toward building AI factories capable of supporting massive inference and training workloads, infrastructure must evolve from standalone systems to integrated rack-scale designs.\n\nThat shift took center stage at Lenovo Tech World ’26, where AMD Chair and CEO Dr. Lisa Su announced Lenovo as one of the earliest OEM adopters of the AMD Helios™ rack-scale solution. Lenovo is applying decades of infrastructure engineering expertise to help customers build the next generation of AI factories – from scale-up and scale-out cluster design to the deployment and lifecycle services that keep AI environments running at scale.\n\nDesigned for the next era of generative and agentic AI, AMD Helios™ brings together next-generation AMD Instinct™ GPUs, AMD EPYC™ CPUs, AMD Pensando™ networking technologies and the AMD ROCm™ software ecosystem into an open, rack-scale architecture built on Open Compute Project (OCP) and Open Rack Wide (ORW) specifications. Engineered for hyperscale inference, training and fine-tuning, the solution is built to scale from rack-level systems into datacenter-scale clusters for distributed inference and the largest foundation-model training runs.\n\nTo support AI at this scale, the platform combines **72 AMD Instinct™ MI455X GPUs**, next-generation AMD EPYC™ processors, **31TB of HBM4 memory**, and AMD Pensando™ networking technologies in a single rack. The result is **up to 2.9 exaFLOPS of FP4 AI inference performance** and **1.4 exaFLOPS of FP8 AI training performance**, giving hyperscalers and NeoCloud providers the performance needed to train, fine-tune, and serve increasingly sophisticated AI models.\n\n“AI infrastructure is rapidly evolving from standalone servers to fully integrated rack-scale clusters,” said Conor Malone, Vice President and General Manager of Cloud Service Providers, Lenovo. “Working with AMD, we’re bringing together world-class cluster design with Lenovo’s expertise in deploying at scale, helping customers leverage this solution to build AI environments that are open, scalable, and ready for what’s next.”\n\n**More Than Hardware: A Faster Path to Deployment**\n\nPreparing organizations for the new era of AI isn’t simply about deploying new hardware. It requires expertise in cluster design, liquid cooling, workload optimization, and lifecycle management. Leveraging its proven experience in deploying large-scale AI infrastructure, Lenovo helps customers accelerate deployment, optimize performance, reduce time to first token, and move AI environments into production faster.\n\nLenovo’s engineering and services experts work alongside customers to assess workload requirements, right-size deployments, optimize GPU configuration, and tune performance for the workload at hand. From there, ongoing full-lifecycle services capabilities reduce stand-up time and support long-term differentiation with greater speed and efficiency, helping organizations maximize GPU utilization and reduce deployment risk.\n\nThe AI factory may be changing how infrastructure is built, but Lenovo’s goal remains the same: help organizations deploy AI where it delivers the greatest value. Through its Hybrid AI vision, Lenovo is building the open, scalable infrastructure that enables customers to run AI across edge, enterprise, cloud, and hyperscale environments—meeting them where they are today while helping them prepare for what’s next. Lenovo solutions based on the AMD Helios rack-scale solution are expected to be available in 4Q 2026.", "url": "https://wpnews.pro/news/the-next-evolution-of-ai-infrastructure-inside-the-architecture-powering-the-ai", "canonical_source": "https://news.lenovo.com/inside-the-architecture-powering-the-ai-factory-era/", "published_at": "2026-07-23 18:30:25+00:00", "updated_at": "2026-07-23 18:53:16.337514+00:00", "lang": "en", "topics": ["artificial-intelligence", "ai-infrastructure", "ai-chips", "ai-products"], "entities": ["AMD", "Lenovo", "Lisa Su", "AMD Helios", "AMD Instinct MI455X", "AMD EPYC", "AMD Pensando", "Conor Malone"], "alternates": {"html": "https://wpnews.pro/news/the-next-evolution-of-ai-infrastructure-inside-the-architecture-powering-the-ai", "markdown": "https://wpnews.pro/news/the-next-evolution-of-ai-infrastructure-inside-the-architecture-powering-the-ai.md", "text": "https://wpnews.pro/news/the-next-evolution-of-ai-infrastructure-inside-the-architecture-powering-the-ai.txt", "jsonld": "https://wpnews.pro/news/the-next-evolution-of-ai-infrastructure-inside-the-architecture-powering-the-ai.jsonld"}}