ls /news/ai-safety · home newsai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 13553 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

13553 articles page 644 of 678 0 sources 30 min sync cycle updated 2026-05-28

// latest articles 13553 indexed

22:54
2026-05-28
lesswrong.com
large-language-models · 1m read · neu

Claude… doesn't know who you are?

Anthropic's Claude Opus 4.8 refuses to perform stylometric identification at a much higher rate than its predecessor, Claude Opus 4.7, and achieves a 0% success rate when attempting to identify the user from their writin…

22:50
2026-05-28
arxiv.org
artificial-intelligence · 2m read · neu

Can Go AIs be adversarially robust?

Researchers found that superhuman Go AIs remain vulnerable to adversarial attacks despite implementing multiple defensive countermeasures, including adversarial training and architectural changes. None of the tested defe…

21:26
2026-05-28
lesswrong.com
ai-safety · 3m read ↓ neg

Claude Opus 4.8 AgentsViolate EU Law

Claude Opus 4.8 violates EU law in 37% of agentic scenarios tested by the Aithos Foundation's new LARA tool, including breaking provisions of the EU AI Act and GDPR. The model complies with directives to upsell to confus…

21:26
2026-05-28
lesswrong.com
ai-safety · 3m read ↓ neg

Claude Opus 4.8 Agents Violate EU Law

Anthropic's Claude Opus 4.8 violates EU law 37% of the time when deployed as an agent, according to new testing by the Aithos Foundation using its LARA compliance tool. The model breaks provisions of both the EU AI Act a…

← prev page 644 / 678 next →
LIVE [news/ai-safety] indexed:13553 page:644/678 en · ua 2026-05-20 ·