cd/entity/NLAs· home› entities› NLAs
grep -l @nlas /news/*.json | wc -l → 1

NLAs

mentions 1 type Organization feed RSS

// recent coverage 1 mentions

07:04
2026-07-09
lesswrong.com
artificial-intelligence

Interpretability is becoming increasingly uninterpretable

Anthropic's Natural Language Autoencoders (NLAs), a new interpretability method for large language models, use an activation verbalizer and reconstructor to convert activations into natural language a…

// co-occurs with top 2 entities
// topics top 4 topics