# MoME: Mixture-of-Memory Embeddings for Context-Aware Sparse Lookup

> Source: <https://aiflash.com/news/123538/>
> Published: 2026-09-21 05:30:01+00:00

Scaling large language models efficiently has motivated sparse capacity mechanisms such as Mixture-of-Experts and, more recently, conditional memory: token-indexed embedding tables that augment the backbone with cheap parametric lookups. Existing memory-embedding methods retrieve via a deterministic
