cd/entity/LongBench· home entities LongBench
grep -l @longbench /news/*.json | wc -l → 9

LongBench

mentions 9 type Organization feed RSS

// recent coverage 9 mentions

07:39
2026-07-11
machinebrief.com
artificial-intelligence

Transformer Efficiency: A Closer Look at KV Cache Compression

New research characterizes the intrinsic compressibility of KV caches in Transformer models, proposing a principled algorithm for efficient inference. The study uses minimax risk to determine when acc…

20:52
2026-06-29
arxiv.org
machine-learning

Simplified Sparse Attention via Gist Tokens

Researchers introduced Simplified Sparse Attention (SSA), a method that uses gist tokens to achieve sparse attention without architectural changes, outperforming baselines on LongBench and improving r…

22:29
2026-06-14
arxiv.org
large-language-models

Still: Amortized KV Cache Compaction in a Single Forward Pass

Researchers introduced Still, a per-layer Perceiver model that compacts KV cache in a single forward pass, enabling efficient long-context language model deployment. On Qwen and Gemma models, Still ou…

// co-occurs with top 8 entities
// topics top 6 topics