cd/entity/Best-of-N· home› entities› Best-of-N
grep -l @best-of-n /news/*.json | wc -l → 3

Best-of-N

mentions 3 type Organization feed RSS

// recent coverage 3 mentions

04:00
2026-10-08
machinebrief.com
artificial-intelligence

Efficient Best-of-N policy evaluation for inference-time alignment

A new arXiv paper (2610.09250v1) proposes a sample-only framework for evaluating and selecting Best-of-N (BoN) inference-time alignment policies without access to response likelihoods, which standard …

04:00
2026-07-07
arxiv.org
large-language-models

Safe Inference-Time Alignment via Lagrangian Reward Augmentation

Researchers propose Lagrangian Reward Augmentation (LARA), a framework for inference-time alignment of language models that enforces safety constraints by dualizing a constrained optimization problem …

// co-occurs with top 5 entities
// topics top 6 topics