00:00
2026-07-16
machinelearning.apple.com
large-language-models
Embarrassingly Simple Self-Distillation Improves Code Generation
Simple self-distillation (SSD) improves LLM code generation by fine-tuning models on their own sampled outputs, boosting Qwen3-30B-Instruct from 42.4% to 55.3% pass@1 on LiveCodeBench v6, with gains cโฆ