cd /news/artificial-intelligence/what-does-chain-of-thought-contribut… · home topics artificial-intelligence article
[ARTICLE · art-76477] src=machinebrief.com ↗ pub= topic=artificial-intelligence verified=true sentiment=· neutral

What Does Chain-of-Thought Contribute at Probe Time? Evidence for Local Co-Occurrence Activation

A new study from arXiv finds that chain-of-thought (CoT) prompting improves large language model performance primarily through local word co-occurrence activation rather than global reasoning order. Researchers observed that randomizing rationale sentence order had little effect on accuracy, while restoring only short-range word order—often with just three-word windows—recovered most of the CoT benefit. The findings suggest that the probe-time advantage of fixed rationales stems from words and short-range co-occurrences, not from the logical sequence of reasoning steps.

read1 min views1 publishedJul 28, 2026

arXiv:2605.26795v2 Announce Type: replace Abstract: Chain-of-thought (CoT) prompting enhances large language model performance, yet what drives these gains remains unclear. We study this question from a probe-time perspective: holding CoT rationales fixed, we test which textual properties matter for the final prediction. Across multiple datasets and model configurations, we find that randomizing the order of rationale sentences has little effect on accuracy, suggesting that the global order of reasoning steps is not the main source of the probe-time benefit. Moreover, even when the words in a rationale are randomly reordered, performance remains well above the no-rationale baseline, indicating that the rationale's words remain useful even without their original order. Restoring only short-range word order further improves performance and brings it substantially closer to full CoT. In most settings, much of this local-order gain is already obtained with three-word windows. Control experiments rule out explicit answer copying, simple lexical cues, generic topical context, and general robustness to shuffling as the main explanations. Mechanistic analyses further show that short-window gains are largely formed in early-to-middle model layers, with answer-relevant evidence concentrated in local text spans. Together, these findings support a local co-occurrence activation (LCA) interpretation: the probe-time benefit of fixed rationales arises mainly from the words they contain and short-range word co-occurrences.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @arxiv 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/what-does-chain-of-t…] indexed:0 read:1min 2026-07-28 ·