02:24
2026-09-03
arxiv.org
large-language-models
Language Models Can Control Their Own Attention
Researchers introduced Declarative Attention (DA), a protocol that lets language models declare which parts of their context they need to attend to, reducing total attended tokens by 52.0% on Gemma-4-…