04:00
2026-07-30
arxiv.org
artificial-intelligence
Post-Training at the Edge of Detectability: A Game-Theoretic Approach to Fine-Tuning
A new game-theoretic framework for reinforcement learning fine-tuning, proposed in arXiv:2607.26358, treats the trade-off between reward and policy drift as a sequential game where a monitor tests forβ¦