cd/entity/Fraser-Taliente· home entities Fraser-Taliente
grep -l @fraser-taliente /news/*.json | wc -l → 2

Fraser-Taliente

mentions 2 type Organization feed RSS

// recent coverage 2 mentions

13:51
2026-07-15
lesswrong.com
ai-safety

Eliciting hidden knowledge from monitors with NLAs

Researchers propose using natural language autoencoders (NLAs) to surface hidden reasoning from AI monitors, testing whether NLAs can recover knowledge of reward hacking that monitors internally detec…

00:43
2026-07-10
lesswrong.com
artificial-intelligence

How robust are natural language autoencoders to initialization?

Researchers at MATS found that natural language autoencoders (NLAs) for LLMs can achieve high reconstruction accuracy even when initialized with entirely implausible statements, emitting 99.3% implaus…

// co-occurs with top 6 entities
// topics top 5 topics