04:00
2026-08-28
arxiv.org
artificial-intelligence
AdaThinking-E: One-Token Entropy Regulation for Adaptive Thinking
Researchers propose AdaThinking-E, a reinforcement learning framework that uses one-token entropy regulation to enable multimodal large language models to adaptively decide when to engage in deep reasβ¦