Unpacking Latent Reasoning: Faithfulness in AI Inference
Researchers studying latent reasoning in AI models found that the causal impact of reasoning steps decays during training, making models less faithful for binary choices but more faithful for open-ended answers. This rai…