Paper Proposes Grading AI Security Agents Without Labels by Measuring Convergence to a Stronger Model
Five researchers posted a preprint on arXiv on 11 August 2026 proposing a label-free method to evaluate agentic continual learning harnesses by measuring how much a smaller student model converges tow…