04:00
2026-07-28
machinebrief.com
large-language-models
LOCKS: Page-Local Compact Key Summaries for Efficient Long-Context Decoding
LOCKS, a new method from arXiv, enables efficient long-context decoding by giving each page of the key-value cache its own compact spectral summary, reconstructing within-page logits, and attending onβ¦