Not All LLM Reasoning is Visible in the Chain-of-Thought
A new study on arXiv (2607.22925v1) shows that frontier language models can perform invisible reasoning using semantically irrelevant filler tokens, with accuracy improvements of up to 13 percentage p…