cd /news/ai-agents/heavy-tailed-memory-traces-in-long-h… · home › topics › ai-agents › article
[ARTICLE · art-143622] src=arxiv.org ↗ pub= topic=ai-agents verified=true sentiment=↑ positive

Heavy-Tailed Memory Traces in Long-Horizon Language Agents

A new arXiv paper (2610.00010v1) proposes Core-Tail World Model (CTWM), a rank-based memory controller that allocates prompt budget with a single exponent τ while retaining a summarized tail, after a tail audit found that semantic LLM policies yield the strongest truncated-power-law-compatible core-tail memory traces. On Synthetic Graph World, CTWM preserved full state and transition coverage, cut prompt tokens by 5.9%, and lowered bottom-half tail prediction error by 13.6% versus a graph-memory baseline, with a 24.48% token reduction on LongMemEval at aggregate accuracy parity. The authors argue heavy-tailed memory traces are both a diagnostic of finite retrieval and a practical control signal for token-efficient long-horizon language agent world models.

by read1 min views4 publishedOct 2, 2026

arXiv:2610.00010v1 Announce Type: new Abstract: Long-horizon language agents increasingly rely on external memory as a frozen world model, yet current memory systems are usually judged only by task success or token cost. We argue that the missing object is the shape of memory use: under finite context and repeated retrieval, agent memory can concentrate on a small core while leaving rare states in a long tail where prediction errors accumulate. We study this effect through a conservative tail audit and find that concentration is reproducible but policy-dependent. Random-walk agents produce log-normal-compatible retrieval artifacts, whereas semantic LLM policies yield the strongest truncated-power-law-compatible core--tail traces. Motivated by this audit, we propose Core--Tail World Model (CTWM), a rank-based memory controller that allocates prompt budget with a single exponent $\tau$ while retaining a summarized tail. On Synthetic Graph World, CTWM preserves full state and transition coverage, reduces prompt tokens by 5.9%, and lowers bottom-half tail prediction error by 13.6% relative to a graph-memory baseline. The same paired comparison gives consistent token savings on ALFWorld and a 24.48% token reduction on LongMemEval with aggregate accuracy parity. These results suggest that heavy-tailed memory traces are not only a diagnostic of finite retrieval, but also a practical control signal for token-efficient agent world models.

── more in #ai-agents 4 stories · sorted by recency
── more on @core-tail world model 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/heavy-tailed-memory-…] indexed:0 read:1min 2026-10-02 · —