04:00
2026-07-22
machinebrief.com
large-language-models
Copy Less, Ground More: Overcoming Repetitive Copying in Long-Context Reasoning via Evidence-Aware Reinforcement Learning
Researchers at arXiv identified a failure mode called repetitive copying in long-context large language models, where models extensively copy input text instead of reasoning. They propose GEAR (Groundβ¦