cd /news/ai-agents/evaluating-long-term-memory-for-ai-a… · home topics ai-agents article
[ARTICLE · art-135642] src=twitter.com ↗ pub= topic=ai-agents verified=true sentiment=↑ positive

Evaluating Long-Term Memory for AI Agents: AML Cycle 2 Is Now Open

The Agent Memory Leaderboard opened Cycle 2 of the Agent Memory Challenge 2026 on September 20, offering over USD 22,000 in prizes to eligible open-source teams. The challenge evaluates long-term agent memory across three tracks — Textual, Coding, and Multimodal — using a shared Add/Search interface and standardized Answer/Eval scoring to test whether agents retrieve current, useful evidence rather than stale context. Both open-source methods and commercial products can compete, with public, comparable results published at agentmemoryleaderboard.ai/evaluation.

read2 min views1 publishedSep 21, 2026
Evaluating Long-Term Memory for AI Agents: AML Cycle 2 Is Now Open
Image: source

Agent Memory Leaderboard on X: "Agent Memory Challenge 2026 Cycle 2 is now open. Long-term memory is not just about storing more history. It is about retrieving the right evidence, recognizing what has changed, and avoiding stale context when an agent needs to act. Three tracks: Textual · Cod… / X

Agent Memory Leaderboard on X: "Agent Memory Challenge 2026 Cycle 2 is now open. Long-term memory is not just about storing more history. It is about retrieving the right evidence, recognizing what has changed, and avoiding stale context when an agent needs to act. Three tracks: Textual · Coding · Multimodal Open-source Methods · Commercial Products Over USD 22,000 prize pool for eligible open-source teams. A shared Add/Search interface. Standardized Answer/Eval. Public, comparable results.

Join: https://t.co/t2QajOb7Fe" Agent Memory Challenge 2026 Cycle 2 is now open. Long-term memory is not just about storing more history. It is about retrieving the right evidence, recognizing what has changed, and avoiding stale context when an agent needs to act. Three tracks: Textual · Coding · Multimodal Open-source Methods · Commercial Products Over USD 22,000 prize pool for eligible open-source teams. A shared Add/Search interface. Standardized Answer/Eval. Public, comparable results. Join: agentmemoryleaderboard.ai/evaluation

Agent Memory Challenge 2026 Cycle 2 is now open. Long-term memory is not just about storing more history. It is about retrieving the right evidence, recognizing what has changed, and avoiding stale context when an agent needs to act. Three tracks: Textual · Coding · Multimodal Open-source Methods · Commercial Products Over USD 22,000 prize pool for eligible open-source teams. A shared Add/Search interface. Standardized Answer/Eval. Public, comparable results. Join: agentmemoryleaderboard.ai/evaluation

For the full technical overview of Cycle 2—including the shared Add/Search evaluation boundary, Textual, Coding, and Multimodal tracks, and the principles behind reproducible Agent Memory evaluation: Agent Memory Challenge 2026 Cycle 2 opens September 20. A shared evaluation for long-term Agent Memory across Textual, Coding, and Multimodal tracks—testing not just what agents store, but whether they retrieve current, useful evidence when it matters.

── more in #ai-agents 4 stories · sorted by recency
── more on @agent memory leaderboard 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/evaluating-long-term…] indexed:0 read:2min 2026-09-21 ·