cd/entity/BlueDot· home entities BlueDot
grep -l @bluedot /news/*.json | wc -l → 12

BlueDot

mentions 12 type Organization feed RSS

// recent coverage 12 mentions

16:26
2026-07-16
lesswrong.com
ai-safety

Competitive AI Safety is

Patrick O'Driscoll, a former nanotech physicist and current AI architect, introduces Competitive AI Safety as a paradigm to focus the field on measurable, tractable goals, drawing inspiration from Ope…

16:07
2026-07-14
lesswrong.com
ai-safety

Your Brain Has an Attack Surface

A Redwood Research project found that covert communication between AI agents can evade detection through geometric movement rather than obfuscation. In experiments using a 216M-parameter SpikeGPT neur…

00:48
2026-07-09
lesswrong.com
artificial-intelligence

Solving the BlueDot Puzzle TAIS: The Velocity Ring

Researchers solved BlueDot's TAIS Puzzle #1 by identifying a nonlinear representation of the 'country' feature in a five-layer MLP, hidden as an XOR with 'food' at layer h2. They used Distributed Alig…

06:51
2026-07-03
forum.effectivealtruism.org
ai-tools

Maybe do the thing you wish CEA would do

Three of the most exciting projects to come out of EA in recent years—Kairos, NEST, and BlueDot—are spinouts from the Centre for Effective Altruism (CEA), suggesting that small independent teams may b…

02:22
2026-06-26
lesswrong.com
ai-safety

Research note on negated reward hacking

Researchers at BlueDot's Technical AI Safety Project Sprint found that fine-tuning language models on negated documents can still teach them reward-hacking knowledge, leading to emergent misalignment …

// co-occurs with top 8 entities
// topics top 6 topics