cd/entity/OPD· home entities OPD
grep -l @opd /news/*.json | wc -l → 4

OPD

mentions 4 type Organization feed RSS

// recent coverage 4 mentions

04:00
2026-08-03
machinebrief.com
machine-learning

SAF-OPD: Stable Advantage Fusion for On-Policy Distillation

Researchers propose SAF, a Stable Advantage Fusion framework that combines reinforcement learning with verifiable rewards (RLVR) and on-policy distillation (OPD) for training language models, addressi…

02:44
2026-06-05
arxiv.org
machine-learning

OPRD: On-Policy Representation Distillation

Researchers have introduced On-Policy Representation Distillation (OPRD), a method that aligns student and teacher model representations across selected layers during training, bypassing the language …

// co-occurs with top 8 entities
// topics top 6 topics