cd /news/ai-chips/qualcomm-to-unveil-snapdragon-8-elit… · home topics ai-chips article
[ARTICLE · art-137486] src=cryptobriefing.com ↗ pub= topic=ai-chips verified=true sentiment=↑ positive

Qualcomm to unveil Snapdragon 8 Elite Gen 6 smartphone chips with AI focus

Qualcomm will fully unveil its Snapdragon 8 Elite Gen 6 smartphone chip at the Snapdragon Summit on September 22 in Maui, with the processor able to run a 30 billion parameter mixture-of-experts AI model entirely on-device. The chip activates only about 3 billion of those parameters per inference step and pairs a revamped Hexagon NPU, which adds a dedicated Element Accelerator for transformer workloads and 50% more shared memory for context windows up to 32K tokens, with custom Oryon CPU cores in a 2+3+3 configuration. Both the standard and Pro variants are expected to use TSMC's 2nm process node, and Qualcomm is positioning the silicon as enabling 'agentic AI' workflows that keep sensitive data off remote servers.

read2 min views3 publishedSep 22, 2026
Qualcomm to unveil Snapdragon 8 Elite Gen 6 smartphone chips with AI focus
Image: Cryptobriefing (auto-discovered)

The new mobile processors can run 30 billion parameter AI models entirely on-device, signaling a major shift toward local inference and away from cloud dependence

Qualcomm just pulled back the curtain on its next generation of flagship mobile silicon, and the headline number is a big one. The Snapdragon 8 Elite Gen 6, set for a full unveiling at the Snapdragon Summit on September 22 in Maui, can run a 30 billion parameter mixture-of-experts AI model locally on a smartphone.

What’s actually under the hood #

The 30 billion parameter figure comes with an important asterisk that makes the whole thing work. The chip uses a Mixture-of-Experts (MoE) architecture, meaning the full model contains 30 billion parameters, but only about 3 billion are activated during any given inference step.

The key enabler is Qualcomm’s revamped Hexagon NPU, which now includes a dedicated “Element Accelerator” built specifically for transformer workloads. The NPU also gets 50% more shared memory compared to the previous generation, which allows context windows up to 32K tokens.

On the CPU side, the chip uses Qualcomm’s custom Oryon cores in a 2+3+3 configuration: two high-performance cores, three balanced cores, and three efficiency cores. The Adreno GPU gets an upgrade too, with new Matrix cores designed for AI-assisted graphics rendering.

AI, tech, and the markets they move—in one daily briefing.

Daily. Free. Join 34,000+ readers across crypto, finance, and policy.

Both the standard and Pro variants are expected to be built on TSMC’s cutting-edge 2nm process node. For the Pro version specifically, Qualcomm is adding extra Matrix ALU blocks to the GPU for what the company calls “AI Frame Fusion” capabilities.

The privacy play #

Beyond raw performance, Qualcomm is leaning hard into a narrative about on-device AI as a privacy advantage. The company frames its new chips as enabling “agentic AI” workflows, where an AI assistant can take multi-step actions on your behalf, without routing sensitive data through remote servers.

Disclosure: This article was edited by Editorial Team. For more information on how we create and review content, see our

Editorial Policy.

── more in #ai-chips 4 stories · sorted by recency
── more on @qualcomm 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/qualcomm-to-unveil-s…] indexed:0 read:2min 2026-09-22 ·