04:00
2026-09-11
machinebrief.com
large-language-models
Negative Self-Distillation: Learning to Reason by Avoiding Flaws
A new arXiv paper (2609.11699v1) introduces Negative Self-Distillation (NSD), a framework that improves large language model reasoning by diverging from self-generated flawed reasoning rather than imi…