Reasoning steps like calculation, formula retrieval, and deduction are clearly separable in a model's internal states, especially in the middle layers. That matters for AI safety, because models process more than their visible chain of thought reveals.
The article AI models' written reasoning steps correspond to distinct internal patterns, a new study finds appeared first on The Decoder.