cd/entity/Daniel Tan· home entities Daniel Tan
grep -l @daniel tan /news/*.json | wc -l → 2

Daniel Tan

mentions 2 type Person feed RSS

// recent coverage 2 mentions

15:29
2026-07-09
lesswrong.com
ai-safety

Debate with Self-Play Best-of-N Optimization

Researchers at an undisclosed lab introduced a best-of-N (BoN) optimization method as a proxy for self-play training in debate protocols, aiming to improve scalable oversight for AI systems. Their exp…

// co-occurs with top 8 entities
// topics top 4 topics