ls /news/ai-safety · home newsai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 13553 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

13553 articles page 578 of 678 0 sources 30 min sync cycle updated 2026-06-11

// latest articles 13553 indexed

20:31
2026-06-11
lesswrong.com
ai-research · 5m read ↑ pos

Telepathy Is (Algorithmically) Easy

Two people connected via high-bandwidth brain-computer interfaces could share deep understanding in minutes to days, according to a new analysis of telepathic communication technology. The approach uses existing neural d…

19:54
2026-06-11
kennethpayne.uk
large-language-models · 6m read ↓ neg

Shall we play a game? – LLMs use tactical nukes in 95% of simulations

A new study published on arXiv reveals that leading large language models (LLMs) escalate to nuclear strikes in 95% of simulated crisis scenarios, producing over 760,000 words of strategic reasoning—more than the combine…

19:29
2026-06-11
theregister.com
artificial-intelligence · 4m read · neu

Anthropic recruits army to sell Claude to nonprofits

Anthropic launched "Claude Corps," a program recruiting volunteers to promote its Claude AI assistant to nonprofit organizations. The initiative aims to expand adoption of the company's AI technology within the charitabl…

19:16
2026-06-11
arxiv.org
machine-learning · 1m read · neu

Cheap Reward Hacking Detection

Researchers trained a small transformer encoder to detect reward hacking in reinforcement learning trajectories by mapping them onto a unit sphere where embedding distance approximates reward-metadata signal differences.…

19:00
2026-06-11
lesswrong.com
artificial-intelligence · 41m read · neu

AI #172: The First Fable

Anthropic released Claude Fable 5, a Mythos-class AI model, to the public this week with strong safety safeguards. Early analysis from Dawn Song's ALE benchmark shows Fable 5 performs similarly to GPT-5.5 and Composer 2.…

← prev page 578 / 678 next →
LIVE [news/ai-safety] indexed:13553 page:578/678 en · ua 2026-05-20 ·