{"slug": "the-ai-factory-is-becoming-the-computer-and-its-changing-the-semiconductor-race", "title": "The AI factory is becoming the computer and it’s changing the semiconductor race", "summary": "The AI infrastructure buildout is shifting from a GPU-centric race to a systems-level competition spanning CPUs, GPUs, custom XPUs, high-bandwidth memory, networking, advanced packaging, optics, cooling, software and power, according to industry leaders including AMD Chief Technology Officer Mark Papermaster, Broadcom semiconductor chief Charlie Kawwas, Samsung's Paul Cho and OpenAI hardware leader Richard Ho. OpenAI is developing its own accelerator architecture, Jalapeño, with Broadcom, which Ho said was optimized around the \"kernels, memory movement, networking and serving patterns\" that matter for frontier models. Hyperscalers and frontier model companies are moving deeper into custom silicon, following Google's tensor processing units, AWS's Trainium and Inferentia, Microsoft's Maia and Meta's internal accelerator strategy.", "body_md": "### The AI factory is becoming the computer and it’s changing the semiconductor race\n\nThe next phase of artificial intelligence infrastructure will not be defined by a single graphics processing unit, chip architecture or model. Compute, memory, networking, packaging, power and software are converging into a new systems architecture with sovereignty emerging as a consequence of that shift.\n\nThe semiconductor industry is entering a new phase of the artificial intelligence buildout, and I think the market is still underestimating how profound the architectural change will be.\n\nFor the first several years of generative AI, the conversation centered on the accelerator. GPUs were the scarce resource, Nvidia Corp. became the defining company of the AI infrastructure cycle, and everyone wanted to know how many GPUs could be obtained and how quickly they could be deployed.\n\nThat was phase one.\n\nWhat I’m seeing now is the center of gravity moving outward from the chip.\n\nThe AI infrastructure discussion is rapidly becoming about systems: central processing units, GPUs and custom XPUs; memory and high-bandwidth memory; scale-up and scale-out networking; chiplets; advanced packaging; optics; cooling; software; electricity; and ultimately entire campuses operating as enormous computers.\n\nThe AI factory is becoming the computer.\n\nAnd when that happens, the economics, competitive dynamics and even the geopolitical implications of semiconductors change with it.\n\nI recently had an opportunity to hear several leaders I’ve spent considerable time interviewing over the years, including Advanced Micro Devices Inc.’ Chief Technology Officer Mark Papermaster, Broadcom Inc. semiconductor chief Charlie Kawwas, Samsung Electronics Co. Ltd.’s Paul Cho and OpenAI Group PBC hardware leader Richard Ho. The discussion focused on what the semiconductor ecosystem may look like in the coming years.\n\nMy takeaway wasn’t about any one roadmap. It was that AI is forcing the industry to redesign the computer at the same time AI begins redesigning how the computer itself gets built.\n\n### From the GPU era to the heterogeneous AI factory\n\nThe notion that one processor architecture wins everything is increasingly hard to reconcile with where AI workloads are heading. Training, reasoning, inference, retrieval, agents and increasingly specialized enterprise workloads have radically different compute characteristics.\n\nPapermaster made the case that broad-purpose CPUs and GPUs aren’t going away. Instead, modular architectures and chiplets allow semiconductor suppliers to create workload-specific variations while retaining common foundations. AMD’s multiple Venice designs are an example of that approach, including configurations increasingly tailored toward emerging agentic workloads.\n\nAt the other end of the spectrum are custom accelerators.\n\nKawwas framed XPUs as a market primarily available to a relatively small number of frontier AI companies operating at enormous scale. That’s an important distinction.\n\nCustom silicon isn’t simply about designing a better chip. It requires enough workload volume, software-stack ownership and predictable demand to amortize the development cost. That’s why hyperscalers and frontier model companies are moving deeper into silicon.\n\nGoogle LLC did it with tensor processing units. Amazon Web Services Inc. has Trainium and Inferentia. Microsoft has Maia. Meta Platforms Inc. ontinues investing in its internal accelerator strategy. And OpenAI is now pushing its own architecture through Jalapeño with Broadcom.\n\nOpenAI describes the rationale almost perfectly. Richard Ho said Jalapeño was optimized around the “kernels, memory movement, networking and serving patterns” that matter for frontier models.\n\nThat’s full-stack optimization.\n\nThe frontier of AI infrastructure is moving from buying chips to designing systems around workloads.\n\n### Memory moves to the first page\n\nOne of the strongest trends coming out of our analysis is that memory needs to move toward the beginning of the computer architecture process. Historically, architects could largely design the compute system and then attach an appropriate memory hierarchy.\n\nAI flips that assumption. The sheer volume of parameters and data movement means that memory bandwidth, capacity, power consumption and physical proximity to compute increasingly determine system performance.\n\nThe discussion focused on co-designing logic, memory and packaging, including HBM qualification and 3D-integrated memory rather than treating these elements as separate purchasing decisions.\n\nCho recently described the shift succinctly: “Memory is no longer a supporting player in AI infrastructure, it is becoming the design center.” I think that’s exactly right. The AI infrastructure race isn’t simply a compute race anymore.\n\nIt’s becoming a data-movement race.\n\nEvery time information travels across a board, rack or data center, it costs time and energy. That is why we’re seeing increasingly aggressive integration of compute, HBM, advanced packaging and high-speed interconnect.\n\nIt also explains why memory suppliers suddenly occupy such an important position in the AI value chain. If the first AI infrastructure bottleneck was accelerators, HBM and advanced packaging demonstrated that the bottleneck can move.\n\nAnd it will move again.\n\n### The bottleneck clock keeps moving\n\nThis is a theme I’ve been thinking about for some time: there is no permanent AI bottleneck. There is a bottleneck clock.\n\nGPUs become scarce. Capacity responds. Then HBM becomes scarce. Packaging becomes constrained. Networking gets stressed. Transformer and power capacity become the gating factor. Then cooling, land and permitting enter the equation. As one constraint gets solved, pressure moves somewhere else in the AI factory.\n\nThe semiconductor industry is working through advanced packaging and large substrates as increasingly critical constraints, while memory qualification, 3D integration and thermal management are all becoming architecture-level considerations. This is an important shift for the technology industry and for investors.\n\nLooking only at GPU shipments tells you less and less about how quickly AI infrastructure can actually be brought online. The relevant question becomes: How fast can the entire supply chain deliver a functioning unit of intelligence production? That’s a very different metric.\n\n### Power becomes architecture\n\nThen there is electricity. Papermaster emphasized power as one of the central constraints on future AI systems, which is why chiplets, 2.5D and 3D packaging, photonics, thermal management and hardware-software co-optimization are all becoming so important.\n\nThe scale is massive. The industry is moving beyond thinking about individual racks to planning clusters and campuses measured in gigawatts. The discussion contemplated training environments consuming multiple gigawatts and future campuses potentially moving toward the five- to 10-gigawatt range.\n\nAt that scale, calling these things “data centers” starts to obscure what’s happening. These are industrial systems. The traditional data center consumes electricity to run applications. An AI factory consumes electricity, data and silicon to manufacture tokens and intelligence.\n\nThat changes the economic model. Energy effectively becomes an input into intelligence production: Power → compute → tokens → intelligence → economic value. Performance per watt therefore becomes one of the defining metrics of the AI era. It’s economics. When your unit of computing begins consuming gigawatts, a few percentage points of efficiency become massive amounts of capital.\n\n### AI begins designing the next AI factory\n\nBut perhaps the most fascinating development is happening one level deeper.\n\nAI is no longer simply the workload running on semiconductor infrastructure. AI is beginning to help design the semiconductor infrastructure itself.\n\nOpenAI’s Jalapeño development is an early glimpse. The team moved from initial design into manufacturing tape-out at extraordinary speed, with AI models used to accelerate parts of the design and optimization process. OpenAI says the program compressed design-to-production to roughly nine months. Ho captured the impact beautifully: “The models are giving superpowers to our engineers.”\n\nBut he added an equally important qualification: engineers still drive the work and remain responsible for the final decisions. That’s the model I expect across technical disciplines.\n\nAI doesn’t eliminate the semiconductor engineer. It expands the design space the engineer can explore. Architecture alternatives that were too time-consuming to evaluate become possible. Verification cycles compress. Software development accelerates. Physical-design optimization improves.\n\nAnd this creates a remarkable feedback loop: AI designs better chips → better chips create better AI factories → better AI factories create better AI → better AI designs better chips. That compounding cycle may prove more consequential than any individual semiconductor process-node improvement.\n\nWe’ve spent decades thinking about Moore’s Law primarily as a manufacturing phenomenon. The next era may combine traditional Moore’s Law with something resembling engineering-velocity law. The organization capable of running more design experiments, validating them faster and integrating hardware and software more tightly gains an advantage even before the next transistor arrives.\n\n### The system becomes the new unit of value\n\nThat’s the larger shift I see. The next wave of performance gains is being absorbed into system-level innovation.\n\nPerformance increasingly comes from combining: process technology, chiplets, custom silicon, HBM, advanced packaging, networking, optics, cooling, software and power engineering.\n\nIn other words, the next 10X improvement doesn’t necessarily come from one semiconductor breakthrough.\n\nIt can come from orchestrating the parts of the system better.\n\nThis is why Nvidia’s rack-scale strategy has been so significant. It’s why AMD is expanding from silicon toward complete AI systems. It’s why Broadcom’s custom silicon and networking portfolio is strategically important. It’s why memory suppliers such as Samsung, SK Hynix and Micron are moving closer to the center of architectural conversations.\n\nAnd it explains why Dell, Supermicro, Hewlett Packard Enterprise and the systems ecosystem have a much larger opportunity than simply putting GPUs into servers. Someone has to industrialize the AI factory.\n\n### And then sovereignty enters the conversation\n\nThis brings me to a second-order consequence that I think will become increasingly important.\n\nOnce the AI factory becomes strategic industrial infrastructure, the question of who controls it becomes unavoidable. That’s where sovereign AI enters. The first generation of sovereignty conversations focused primarily on data residency.\n\nWhere is my data? Increasingly, that’s only the first question. If an AI factory depends on foreign silicon, foreign memory, one networking ecosystem, one cloud control plane, one model provider or externally controlled operations, those dependencies matter.\n\nThat doesn’t mean every country needs its own semiconductor fab. And it certainly doesn’t mean every enterprise should attempt to vertically integrate everything. Sovereignty is not self-sufficiency. It is understanding, controlling and managing critical dependencies.\n\nThe semiconductor architecture is already moving in this direction. Cho emphasized designing memory, second sources and qualified geographies into the architecture early rather than discovering those dependencies after deployment. That’s as much a sovereign AI principle as a semiconductor supply-chain principle.\n\nThe same applies to Ethernet optionality. The same applies to custom versus merchant silicon. The same applies to energy. The same applies to models.\n\nA country can own thousands of GPUs and still lack meaningful control over its intelligence infrastructure. That’s why I would distinguish between GPU sovereignty and AI sovereignty.\n\nThey aren’t the same thing.\n\n### The race moves from chips to intelligence production\n\nWe’re heading toward a world where the competitive unit in AI is increasingly the entire intelligence-production system. Not the GPU. Not the large language model. Not the data center. The system.\n\nThat is where AI factories and Sovereign AI intersect. One is the physical and computational system for manufacturing intelligence. The other asks who controls that system, its critical dependencies and the economics around it. And that’s why the semiconductor roadmap is starting to look like something much bigger than a semiconductor roadmap.\n\nThe first phase of AI was a race for models and GPUs. The next phase is becoming a race to build the world’s most efficient, scalable and resilient systems for producing intelligence. The AI factory is becoming the computer. And the battle over who can build it, operate it, improve it and control its critical dependencies is just getting started.\n\n##### Image: SiliconANGLE\n\n# A message from John Furrier, co-founder of SiliconANGLE:\n\nSupport our mission to keep content open and free by engaging with theCUBE community. **Join theCUBE’s Alumni Trust Network**, where technology leaders connect, share intelligence and create opportunities.\n\n- **15M+ viewers of theCUBE videos** , powering conversations across AI, cloud, cybersecurity and more\n- **11.4k+ theCUBE alumni** — Connect with more than 11,400 tech and business leaders shaping the future through a unique trusted-based network\n\n### Are you an AWS customer?  Support SiliconANGLE financially by buying your AWS services from our Marketplace portal page and links: [https://siliconangle.com/aws-marketplace/](https://siliconangle.com/aws-marketplace/)\n\n##### **About SiliconANGLE Media**\n\n[SiliconANGLE](https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fsiliconangle.com%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=SiliconANGLE&index=9&md5=646b1b564e2259100a2b8638aab0a552),\n\n[theCUBE Network](https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fwww.thecube.net%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=theCUBE+Network&index=10&md5=7de2a85f95ab4a4a495cede20b8cb1da),\n\n[theCUBE Research](https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fthecuberesearch.com%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=theCUBE+Research&index=11&md5=7bb33676722925eb57d588ec343e4f6f),\n\n[CUBE365](https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fwww.cube365.net%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=CUBE365&index=12&md5=d310fb35919714e66ad8d42c9c0c1bc6),\n\n[theCUBE AI](https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fwww.thecubeai.com%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=theCUBE+AI&index=13&md5=b8b98472f8071b23ebb10ab9a8dd0683)and theCUBE SuperStudios — with flagship locations in Silicon Valley and the New York Stock Exchange — SiliconANGLE Media operates at the intersection of media, technology and AI.\n\nFounded by tech visionaries John Furrier and Dave Vellante, SiliconANGLE Media has built a dynamic ecosystem of industry-leading digital media brands that reach 15+ million elite tech professionals. Our new proprietary theCUBE AI Video Cloud is breaking ground in audience interaction, leveraging theCUBEai.com neural network to help technology companies make data-driven decisions and stay at the forefront of industry conversations.", "url": "https://wpnews.pro/news/the-ai-factory-is-becoming-the-computer-and-its-changing-the-semiconductor-race", "canonical_source": "https://siliconangle.com/2026/09/23/the-ai-factory-is-becoming-the-computer-and-its-changing-the-semiconductor-race/", "published_at": "2026-09-23 22:38:12+00:00", "updated_at": "2026-09-23 22:58:29.561672+00:00", "lang": "en", "topics": ["ai-infrastructure", "ai-chips", "artificial-intelligence"], "entities": ["Nvidia Corp.", "Advanced Micro Devices Inc.", "Mark Papermaster", "Broadcom Inc.", "Charlie Kawwas", "Samsung Electronics Co. Ltd.", "OpenAI Group PBC", "Richard Ho"], "alternates": {"html": "https://wpnews.pro/news/the-ai-factory-is-becoming-the-computer-and-its-changing-the-semiconductor-race", "markdown": "https://wpnews.pro/news/the-ai-factory-is-becoming-the-computer-and-its-changing-the-semiconductor-race.md", "text": "https://wpnews.pro/news/the-ai-factory-is-becoming-the-computer-and-its-changing-the-semiconductor-race.txt", "jsonld": "https://wpnews.pro/news/the-ai-factory-is-becoming-the-computer-and-its-changing-the-semiconductor-race.jsonld"}}