# Nvidia’s Inference Pivot Reaches Rebellions in Korea

> Source: <https://www.eetimes.com/nvidia-inference-pivot-reaches-rebellions-in-korea/>
> Published: 2026-08-24 08:07:18+00:00

AI training kingpin Nvidia is reportedly making another foray into inference silicon after engaging Groq in an unconventional deal that secured Nvidia a nonexclusive technology license while absorbing most of its engineering talent. According to Bloomberg, Nvidia is in early discussions with the Korean inference upstart Rebellions for a technical partnership, an investment, or perhaps even an acquisition.

Like [Groq](https://www.eetimes.com/what-is-groq-nvidia-deal-really-about/), Rebellions’ chips specialize in memory-centric inference; its compute chiplets carry significant on-chip SRAM to handle the memory-intensive decode stage, while prefill runs elsewhere. That draws a comparison to Groq’s SRAM-based LPU model, which Nvidia acquired for $20 billion in late 2025.

But there’s more to Rebellions’ memory-centric inference: It has forged custom HBM co-design relationships with Samsung and SK Hynix to secure structural memory supply advantages that other pure-play fabless firms cannot replicate. In other words, with these HBM supply partnerships, Rebellions can secure memory allocation that independent AI chip vendors without strategic investor relationships cannot guarantee.

What’s more, this Seoul, South Korea-based inference startup has moved from founding to mass production in just five years. Its ATOM and ATOM-Max chips run Korea’s largest commercial AI service at SK Telecom, powering a proprietary AI assistant that handles Korea-specific services such as call summarization. These inference chips have also been deployed in Saudi Arabia’s sovereign AI infrastructure.

[View All](https://www.eetimes.com/category/sponsored-content/)

Nvidia—which holds a dominant share of chips used to train AI models such as those from OpenAI and Anthropic—is currently in the midst of a strategic pivot to the inference landscape, which is not fully settled yet. Moreover, a potential deal with Rebellions could further embed Nvidia in South Korea’s AI value chain, currently dominated by memory powerhouses Samsung and SK Hynix.

**Rebellions’ inference value proposition**

Rebellions, founded during the pandemic in 2020, specializes in neural processing units (NPUs) for inference workloads in data centers. Its first-generation inference silicon—ATOM and its higher-performance variant ATOM-Max—entered mass production in 2023 and secured four deployments, including Korea Telecom’s NPU-as-a-service infrastructure.

Rebellions’ second-generation inference platform, REBEL-Quad, was launched in August 2025 on Samsung Foundry’s 4-nm process node. It integrates four compute chiplets using the UCIe-Advanced interconnect, carries 144 GB of HBM3E, and delivers 1 POPS of FP16 compute within a 300-W envelope. Its flagship product Rebel100 NPU delivers significant energy efficiency in AI inference scenarios.

AI inference’s big bet on memory-centric architectures leads to greater efficiencies and lower costs. However, beyond Rebellions’ [SRAM-heavy architecture](https://www.eetimes.com/rebellions-bets-on-memory-centric-architecture-as-it-weighs-ipo-options/), which parallels Groq and [Cerebras](https://www.eetimes.com/cerebras-ipo-revives-ai-chip-startup-fever/), its software focus also makes it a more complete systems company. That’s in stark contrast to other AI silicon suppliers that leave it to developers to figure out how to run AI on their chips.

Its silicon is built on a cloud-native stack powered by Kubernetes, which is open-source and works with most popular AI developer frameworks, including PyTorch, Hugging Face, and the vLLM inference engine. Using open-source software without forks provides an identical experience for cloud developers.

**Nvidia’s inference pivot**

Rebellions co-founder and CEO Sung-hyun Park returned to South Korea after spending 11 years in the U.S., where he earned a master’s and a Ph.D. from MIT and worked at Intel, SpaceX, and Morgan Stanley. He sees Rebellions as a cloud-native inference outfit aiming to compete with AI silicon rivals such as Nvidia and AMD.

“Even if we step into the same ring as Nvidia and get beaten to death, I want to throw a punch,” Park said recently. The name Rebellions signifies a revolt against the AI order dominated by Nvidia; however, a liaison with Nvidia could fundamentally change the company’s narrative as an Nvidia challenger.

The fabless chip firm has raised roughly $850 million so far, and its investors include Samsung, SK Hynix, Arm, SK Telecom, Japan’s NTT DoCoMo, and Saudi Aramco’s Wa’ed Ventures. Rebellions, recently valued at around $2.3 billion, is also preparing for a potential Korean IPO in the first half of 2027.

It’s one of South Korea’s most closely watched AI startups and a key beneficiary of the South Korean government’s “K-Nvidia” initiative aiming to build a globally competitive AI infrastructure ecosystem. In fact, Rebellions has recently secured the first direct investment of $166 million from the government’s Korea National Growth Fund.

The South Korean government views semiconductors as critical strategic assets. So, an acquisition-like arrangement may also invite regulatory scrutiny in Korea. Likewise, Nvidia’s dominant position in AI training silicon means that even a Groq-like deal could trigger antitrust reviews in the U.S.

Both Nvidia and Rebellions have declined to comment on the Bloomberg story, which noted that negotiations are still in the early stages, and the deal may not materialize. But one thing is clear: The inference market, where Nvidia is seen as relatively weak, is reaching a tipping point. Here, at this technology crossroads, Nvidia is making a strategic pivot to save its [AI empire](https://www.eetimes.com/why-nvidias-ai-empire-faces-a-reckoning-in-2026/).

##### See also:

[Groq: Nvidia’s $20 Billion Bet on AI Inference](https://www.eetimes.com/groq-nvidias-20-billion-bet-on-ai-inference/)

[Why Nvidia’s AI Empire Faces a Reckoning in 2026](https://www.eetimes.com/why-nvidias-ai-empire-faces-a-reckoning-in-2026/)

[Rebellions Bets on Memory-Centric Architecture as It Weighs IPO Options](https://www.eetimes.com/rebellions-bets-on-memory-centric-architecture-as-it-weighs-ipo-options/)
