{"slug": "case-causal-alignment-and-structural-enforcement-for-improving-chain-of-thought", "title": "CASE: Causal Alignment and Structural Enforcement for Improving Chain-of-Thought Faithfulness", "summary": "Researchers propose CASE, a framework combining training-time causal alignment and inference-time structural enforcement to improve chain-of-thought faithfulness in large language models. Experiments on three models and four benchmarks show CASE achieves a 37% average per-setting relative improvement in overall CoT faithfulness over the strongest baselines. The framework uses counterfactual-CoT, biased-instruction, and empty-instruction datasets with selective-loss fine-tuning, and masks direct attention from instruction tokens to answer tokens during inference.", "body_md": "arXiv:2607.18820v1 Announce Type: new\nAbstract: Chain-of-thought (CoT) reasoning is widely used to improve both the performance and interpretability of large language models (LLMs), yet the generated reasoning may not faithfully support the final answer. We study this problem from a causal perspective, where a faithful CoT process should follow the chain $Z\\rightarrow X\\rightarrow Y$, with $Z$, $X$, and $Y$ denoting the instruction, reasoning chain, and final answer, respectively. In this process, the instruction should affect the answer only through the reasoning chain. However, conventional autoregressive LLMs condition answer generation on both the instruction and the CoT, which still allows a direct instruction-to-answer shortcut. To address this issue, we propose CASE, a framework that combines training-time causal alignment and inference-time structural enforcement. During training, CASE builds counterfactual-CoT, biased-instruction, and empty-instruction datasets, and applies selective-loss fine-tuning to strengthen CoT-to-answer dependence while suppressing instruction shortcuts. During inference, CASE masks direct attention from instruction tokens to answer tokens, preventing the model from bypassing the generated CoT. We provide an information-theoretic analysis showing how these components promote faithful chains. Experiments on three models and four benchmarks show that CASE achieves a 37\\% average per-setting relative improvement in overall CoT faithfulness over the strongest baselines, exhibits stronger cross-dataset faithfulness transfer, and maintains competitive average accuracy. Code is available at https://github.com/oddwang/CASE.", "url": "https://wpnews.pro/news/case-causal-alignment-and-structural-enforcement-for-improving-chain-of-thought", "canonical_source": "https://www.machinebrief.com/news/case-causal-alignment-and-structural-enforcement-for-improvi-2sqi", "published_at": "2026-07-22 04:00:00+00:00", "updated_at": "2026-07-22 04:09:26.232035+00:00", "lang": "en", "topics": ["large-language-models", "ai-research", "ai-safety"], "entities": ["CASE", "arXiv"], "alternates": {"html": "https://wpnews.pro/news/case-causal-alignment-and-structural-enforcement-for-improving-chain-of-thought", "markdown": "https://wpnews.pro/news/case-causal-alignment-and-structural-enforcement-for-improving-chain-of-thought.md", "text": "https://wpnews.pro/news/case-causal-alignment-and-structural-enforcement-for-improving-chain-of-thought.txt", "jsonld": "https://wpnews.pro/news/case-causal-alignment-and-structural-enforcement-for-improving-chain-of-thought.jsonld"}}