Enhancing Assessment of Self-Consistency in LLM Explanations using Perturbation Strength
A new arXiv paper (2609.30849v1) proposes an LLM-as-a-judge approach to measure perturbation strength uniformly across input and chain-of-thought (CoT) perturbations, enabling controlled-strength eval…