cd /news/artificial-intelligence/mesa-task-adaptive-multi-structure-e… · home topics artificial-intelligence article
[ARTICLE · art-93054] src=arxiv.org ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

MESA:Task-Adaptive Multi-Structure Evidence Selection for Long-Horizon Agent Memory

Researchers propose MESA, a task-adaptive multi-structure evidence selection framework for long-horizon agent memory, which dynamically selects and fuses a query-specific subset from five complementary structure views of each trajectory. On AMA-Bench, MESA outperforms the strongest baseline by 8.5% while using 41% fewer evidence tokens than the all-structure alternative, addressing the inefficiency of reading fixed structures or routing to a single one.

read1 min views1 publishedAug 12, 2026

arXiv:2608.10108v1 Announce Type: new Abstract: Long-horizon agents accumulate trajectories spanning hundreds of interleaved reasoning, action, and observation steps, where answering a query may depend on evidence buried far back in the history. External memory stores such trajectories as structured representations, yet each structure provides a distinct and incomplete view. Existing multi-memory systems either read a fixed set of structures for every query, inflating context and introducing noise, or route each query to a single structure, preventing the composition of complementary evidence. A controlled analysis on AMA-Bench shows that the optimal memory configuration is typically neither a single structure nor the full union, but a tailored composition of multiple structural memories that varies with query and task demands. Motivated by these findings, we formulate structure-level dynamic selection: selecting and fusing a query-adaptive subset from a library of specialized memory structures. We propose MESA (a Multi-structure Evidence Selection framework for long-horizon Agent), which builds five complementary structure views of each trajectory and learns from end-to-end answer-level feedback to select and fuse a query-specific subset for a frozen answer model. To learn under this weak supervision, MESA employs harness optimization with prior-guided search and UCB-guided scheduling to balance exploration and exploitation. On AMA-Bench, MESA outperforms the strongest baseline by 8.5% while using 41% fewer evidence tokens than the all-structure alternative.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @mesa 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/mesa-task-adaptive-m…] indexed:0 read:1min 2026-08-12 ·