04:00
2026-07-30
machinebrief.com
large-language-models
OptimismBench: Forecasting Bias and the Alignment Effect in Language Model Judgment
A new study from arXiv introduces OptimismBench, a benchmark that detects directional bias in large language models' probability judgments by using inverted pairs to measure asymmetry between P(succesβ¦