cd /news/artificial-intelligence/shape-of-chain-of-thought-in-math-re… · home topics artificial-intelligence article
[ARTICLE · art-117354] src=arxiv.org ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

SHAPE of Chain-of-Thought in Math Reasoning

Researchers introduced SHAPE, a framework analyzing Chain-of-Thought trajectories in large language models through semantic spaces and heuristics, finding that mathematical heuristics better explain answer correctness than traditional CoT features and that models concentrate reasoning effort in few semantic spaces. The team also showed that reinforcement learning induces mode-seeking in heuristic usage and that post-training with diverse heuristics improves accuracy, with code available on GitHub.

read1 min views1 publishedSep 1, 2026

arXiv:2608.28600v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong performance on mathematical reasoning benchmarks, yet the mathematically meaningful skills underlying their reasoning remain underexplored. We introduce \texttt{SHAPE}, a framework that analyzes Chain-of-Thought (CoT) trajectories through two lenses developed in mathematics education: (1) semantic spaces: the model's evolving mathematical interpretations of a problem (e.g., algebraic, geometric), and (2) heuristics: the specific mathematical actions taken within those spaces (e.g., simplifying the problem, working backward). We first use \texttt{SHAPE} to analyze the reasoning patterns of various models. Our findings reveal that the mathematical heuristics employed by a model better explain final answer correctness than traditional CoT features. Furthermore, models are likely to reach correct solutions by concentrating their reasoning effort within a few semantic spaces rather than exploring many disparate ones -- a pattern consistent with human behavior. Next, we utilize the \texttt{SHAPE} lens to evaluate whether post-training truly enhances mathematical proficiency. We find that reinforcement learning induces mode-seeking in heuristic usage. Lastly, we post-train LLMs by promoting diverse heuristics and demonstrate its effectiveness in improving accuracy. Overall, \texttt{SHAPE} provides a theoretically-grounded diagnostic framework for decoding LLM reasoning and offers a new path toward post-training LLMs for math reasoning. The code for our model is available at https://github.com/holi-lab/SHAPE-of-CoT

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @shape 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/shape-of-chain-of-th…] indexed:0 read:1min 2026-09-01 ·