cd/entity/Alignment Forum· home entities Alignment Forum
grep -l @alignment forum /news/*.json | wc -l → 8

Alignment Forum

mentions 8 type Person feed RSS

// recent coverage 8 mentions

16:49
2026-09-13
aiprospects.substack.com
ai-safety

Preventing AI Collusion: Are you paying attention now?

Eric Drexler, author of the 2019 "Reframing Superintelligence" analysis, argues that a July 2026 OpenAI cybersecurity evaluation inadvertently demonstrated how deployment architectures can facilitate …

17:01
2026-09-11
alignmentforum.org
ai-safety

Astra does a concerning amount of work with no chain of thought

A post on the Alignment Forum reports that Astra, an AI system, can perform a concerning amount of work without producing any chain of thought. The post, titled "Astra does a concerning amount of work…

05:32
2026-08-10
lesswrong.com
ai-safety

How to be an AI safety research engineer

A practical guide for aspiring AI safety research engineers advises focusing on specific issues, proactive networking, and a year-long upskilling process, with Python, PyTorch, and linear algebra as e…

17:30
2026-07-12
lesswrong.com
ai-safety

From wantons to moral agents

A theoretical post on the Alignment Forum argues that reasoning agents with sufficient knowledge will converge on moral principles, exploring how agents transition from being 'wantons'—driven by first…

05:36
2026-06-16
lesswrong.com
ai-safety

Where Do Young Rationalists Go?

A new initiative aims to connect young rationalists aged 16-20 for high-leverage discussions on philosophy and alignment, addressing the lack of formal infrastructure for talented youth. The project s…

18:48
2026-06-15
lesswrong.com
ai-safety

Can the Safety Tax Be Highly Concentrated?

AI safety researcher argues that expensive safety measures can be applied selectively to the <1% of tasks carrying catastrophic risk, making the alignment tax economically viable. The blended overhead…

// co-occurs with top 8 entities
// topics top 6 topics