cd/entity/Olmo 3· home entities Olmo 3
grep -l @olmo 3 /news/*.json | wc -l → 3

Olmo 3

mentions 3 type Person feed RSS

// recent coverage 3 mentions

10:15
2026-07-14
lesswrong.com
ai-safety

Open Distillation of Hereditary Traits

Distilling from Google's Gemma 3 27B IT model into a smaller student model transfers depressive traits, with the student scoring a mean depression of 0.68 on the Gemma Needs Help eval even after aggre…

16:11
2026-06-25
huggingface.co
large-language-models

Which tokens does a hybrid model predict better?

Researchers at the Allen Institute for AI compared their 7B transformer model Olmo 3 with the hybrid model Olmo Hybrid to determine which tokens each predicts better. The hybrid model excels on meanin…

19:45
2026-06-14
lesswrong.com
ai-safety

Why Do Naive SFT Filters For Safety Properties Fail?

Google DeepMind researchers investigate why filtering supervised fine-tuning (SFT) data fails to remove safety-relevant properties from language models, proposing a method to identify the source of th…

// co-occurs with top 8 entities
// topics top 5 topics