cd /news/large-language-models/xiaomis-mimo-v3-to-adopt-new-archite… · home topics large-language-models article
[ARTICLE · art-138882] src=technode.com ↗ pub= topic=large-language-models verified=true sentiment=↑ positive

Xiaomi’s MiMo-V3 to adopt new architecture as HySparse2 cuts long-context costs

Xiaomi MiMo lead Fuli Luo said the upcoming MiMo-V3 model will adopt a new architecture built on HySparse2, a technique that at a context length of 1 million tokens reduces prefill computation by 5.02 times and cuts the KV cache by 4.5 times. Xiaomi reported higher MRCRv2 and RULER-v2 scores along with lower AgentPPL and LongPPL results, positioning HySparse2 as a way to lower long-context inference costs and improve retrieval for agentic workloads.

read1 min views3 publishedSep 24, 2026
Xiaomi’s MiMo-V3 to adopt new architecture as HySparse2 cuts long-context costs
Image: Technode (auto-discovered)

Xiaomi MiMo lead Fuli Luo said MiMo-V3 will adopt a new architecture. Its core technology, HySparse2, is designed to reduce the cost of long-context inference and improve retrieval for agentic workloads.

At a context length of 1 million tokens, HySparse2 reduces prefill computation by 5.02 times and cuts the KV cache by 4.5 times. Xiaomi also reported higher MRCRv2 and RULER-v2 scores, along with lower AgentPPL and LongPPL results. [arXiv]

── more in #large-language-models 4 stories · sorted by recency
── more on @xiaomi 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/xiaomis-mimo-v3-to-a…] indexed:0 read:1min 2026-09-24 ·