cd/entity/Among Usยท homeโ€บ entitiesโ€บ Among Us
grep -l @among us /news/*.json | wc -l โ†’ 1

Among Us

mentions 1 type Person feed RSS

// recent coverage 1 mentions

20:48
2026-07-20
lesswrong.com
ai-safety

Restoring Model Alignment via Honesty Activation Steering

Researchers demonstrate that honesty activation steering can restore model alignment in large language models, with selective steering methods StTP and StMP recovering honesty at a fraction of the capโ€ฆ

// co-occurs with top 5 entities
// topics top 3 topics