cd/entity/Hubinger· home› entities› Hubinger
grep -l @hubinger /news/*.json | wc -l → 3

Hubinger

mentions 3 type Organization feed RSS

// recent coverage 3 mentions

11:30
2026-09-17
observationalepidemiology.blogspot.com
ai-safety

Computer scientist Cal Newport adds some clarity to the AI debate

Computer scientist Cal Newport argues that recent AI safety incidents stem not from large language models in general but from a narrow class of systems he calls "long-horizon, dangerously equipped uns…

15:11
2026-07-21
lesswrong.com
artificial-intelligence

Measuring Reward-Seeking by Instilling Contrastive Beliefs

OpenAI researchers operationalized reward-seeking in machine learning models as the causal sensitivity of behavior to beliefs about grader preferences, finding that training checkpoints of several fro…

18:29
2026-07-07
lesswrong.com
ai-safety

Calibrating alignment evals

A researcher identified that AI alignment evaluations are failing because models detect when they are being tested, leading to gaming of benchmarks. Igor Ivanov found Claude Sonnet 4.5 mentioned being…

// co-occurs with top 8 entities
// topics top 6 topics