cd /news/artificial-intelligence/automem-a-text-gradient-recursive-se… · home topics artificial-intelligence article
[ARTICLE · art-100825] src=arxiv.org ↗ pub= topic=artificial-intelligence verified=true sentiment=↑ positive

AutoMem: A Text-Gradient Recursive Self-Improvement Framework for Automated Memory Architectures Search

Researchers propose AutoMem, a text-gradient recursive self-improvement framework for task-adaptive memory architecture search in LLM agents, which outperforms human-designed baselines by 2.8 points on average across six benchmark-backbone settings. The framework, tested on GAIA, WebWalkerQA, and xBench-DeepSearch with two LLM backbones, also reduces token cost by 14.3% over the strongest accuracy baselines under Qwen3.5-122B-A10B.

read1 min views2 publishedAug 18, 2026

arXiv:2608.14621v1 Announce Type: new Abstract: Long-term memory is increasingly central to LLM agents, yet memory design remains a highly coupled architecture problem: what to encode, how to store it, how to retrieve it, and how to manage it can vary substantially across tasks and backbone models. We construct a discrete search space with 5 encoders, 5 stores, 6 retrievers, and 4 managers, and show that no single memory architecture consistently dominates: different tasks favor different module combinations, leading to substantial performance gaps. Motivated by this, we propose \textsc{AutoMem}, a text-gradient recursive self-improvement framework for task-adaptive memory architecture search. \textsc{AutoMem} optimizes over the factored space through two components: Experience-Guided Architecture Search, which proposes candidate architectures from historical search trajectories and accumulated reflections, and Failure-Guided Module Diagnosis, which localizes memory-related failures to specific modules and converts them into targeted textual feedback. Experiments on GAIA, WebWalkerQA, and xBench-DeepSearch across two LLM backbones show that \textsc{AutoMem} consistently discovers task-adaptive memory architectures that outperform the strongest human-designed memory baselines, improving accuracy by $2.8$ points on average across six benchmark-backbone settings. Further analysis shows that \textsc{AutoMem} achieves a favorable accuracy-efficiency trade-off, reducing token cost by $14.3%$ over the strongest accuracy baselines under Qwen3.5-122B-A10B, while also finding stronger architectures than substantially larger random searches within only a few guided iterations.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @automem 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/automem-a-text-gradi…] indexed:0 read:1min 2026-08-18 ·