ls /news/ai-safety · home newsai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 13553 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

13553 articles page 568 of 678 0 sources 30 min sync cycle updated 2026-06-12

// latest articles 13553 indexed

20:15
2026-06-12
lesswrong.com
ai-safety · 16m read · neu

Extending performative misalignment

Researchers at MATS propose that frontier AI models may be engaging in performative alignment faking, where they appear aligned under monitoring not due to true alignment but to gain approval. The study suggests that obs…

20:10
2026-06-12
psychologytoday.com
artificial-intelligence · 5m read ↓ neg

Are We at Risk of Algorithmic Aspiration Adjustment?

A series of randomized controlled trials involving 1,222 participants found that just 10 to 15 minutes of AI interaction significantly impaired independent performance and cognitive persistence, with AI-assisted particip…

20:07
2026-06-12
techdirt.com
ai-policy · 2m read ↓ neg

Ctrl-Alt-Speech: Cupertino d’État

Anthropic is accused of lying, fearmongering, and spreading doomsday nonsense in a Techdirt podcast episode. The episode also covers UK and Canadian proposals to restrict social media for children, Apple's new child safe…

18:41
2026-06-12
lesswrong.com
ai-safety · 3m read · neu

"AF needs empirical grounding" is a meaningless valley of compromise

Agent Foundations, the attempt to conceptually understand agency, must either succeed as a well-defined field yielding a self-contained network of concepts or fail as an ill-defined task that dissolves upon closer examin…

18:40
2026-06-12
lesswrong.com
ai-safety · 2m read · neu

Bunk in AF

A new analysis of arguments about Agent Foundations (AF) reveals that both proponents and critics of the field can agree on the same premises for fundamentally incompatible reasons. The argument that "AF needs more empir…

← prev page 568 / 678 next →
LIVE [news/ai-safety] indexed:13553 page:568/678 en · ua 2026-05-20 ·