16:10
2026-08-01
lesswrong.com
artificial-intelligence
Do your capabilities homework
A technical AI safety researcher argues that safety-focused researchers should engage with capabilities research, highlighting On-Policy Self-Distillation (OPSD) as a promising alternative to GRPO forβ¦