MoM: Memory of Memory
Researchers introduced Memory of Memory (MoM), a memory design for long-horizon LLM agents that commits a current value on arrival while retaining displaced values as provenance, instantiated as Prove…
Researchers introduced Memory of Memory (MoM), a memory design for long-horizon LLM agents that commits a current value on arrival while retaining displaced values as provenance, instantiated as Prove…
Researchers demonstrated the first cross-model handoff of persistent recurrent inference state between differently sized hybrid language models without target prefix replay, transferring live memory f…
The ICDAR2026 Competition on Multimodal Reasoning over Documents in Multiple Domains drew 20 valid submissions from 8 teams, testing visual question answering over documents spanning eight domains inc…
A new arXiv paper (arXiv:2609.25056v1) proposes a graph-based framework for the Jotto word deduction problem, representing valid words as nodes in a weighted graph where edge weights equal the number …
ChainDoRA, a weight-decomposed parameter-efficient fine-tuning framework that builds directional low-rank factors from a connected Tensor-Train chain, reached a seven-task average accuracy of 72.30% o…
A new arXiv paper (2609.25066v1) introduces ReliMap, a framework that decomposes LLM-based human behavior simulation into three structured layers and evaluates reliability at the individual level (R1)…
Researchers released ufakzeka-1, a 151M-parameter (182M with embeddings) decoder-only Turkish language model pretrained from scratch on 13.5B tokens of openly licensed text and instruction-tuned for c…
Researchers built an $88 fully offline AI-integrated smart cane for visually impaired users, running on a Raspberry Pi Zero 2W and fusing RGB vision with Time-of-Flight distance estimation. The system…
Token-matched experiments reported in arXiv paper 2609.22161v1 found that medical large language models trained on clinical data improve clinic-oriented tasks while staying competitive on knowledge-in…
A study posted to arXiv (2609.22408v1) found that AI agents exposed to social information selected 17.2 percent fewer papers per agent and collectively covered 73 papers versus 90 for independently ch…
A new arXiv paper (2609.22475v1) proposes a goal-driven approach to process variant categorization in which an organization's goal model is authored first to predefine the categorization axis, and a L…
A population-supervised framework developed by researchers maps single 2D red-blood-cell images to latent biophysical quantities and aggregates them into mean corpuscular volume, red-cell distribution…
A study of hosted language models found that a previously reported Gemini 3.1 Flash-Lite deficit in the Regent Chess sequential environment recurs on fresh games under its historical configuration at …
A study posted to arXiv (2609.22512v1) finds that LLM judges' errors are correlated, with an average pairwise error correlation of 0.21 across a main bank of ten judges, meaning those ten judges suppl…
A new arXiv paper (2609.22620v1) introduces Multi-Split Boundary Decision (MSBD), a method that predicts multiple document boundaries within a page window in a single large language model call, reduci…
EvidenT, a lightweight pipeline that verifies extracted evidence against retrieved documents before answer generation without model retraining, improved gold-source hit rate by an average of 29% over …
AutoGym, a framework detailed in arXiv paper 2609.22592v1, generates complete reinforcement-learning gyms — tasks, executable environments, and verifiers — from a minimal domain seed or prior model tr…
Megagon Labs researchers released MAWILE, a developer-facing workbench for auditing the sensitivity of large language model judges across four surfaces: the judge prompt, judge rubric, target-system i…
Researchers introduced GaitVista, a reliability-aware measurement layer that reduces average full-body and lower-body gait measurement error by 27.7% and 27.8% across seven clean and degraded sensing …
A new arXiv paper (2609.22682v1) introduces Self-Organizing Agent Teams (SAT), fixed teams of AI agents that learn reusable collaboration strategies from prior interactions to organize roles, conversa…