ls /news/ai-safety · home newsai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 11073 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

11073 articles page 332 of 554 0 sources 30 min sync cycle updated 2026-06-21

// latest articles 11073 indexed

17:09
2026-06-21
web.archive.org
ai-safety · 3m read · neu

Anthropic uses Persona for identity verification

Anthropic has partnered with Persona to roll out identity verification for certain platform capabilities, requiring users to submit a government-issued photo ID and a live selfie to prevent abuse and comply with legal ob…

16:38
2026-06-21
lesswrong.com
ai-safety · 6m read ↓ neg

How persona training could fail

A scenario warns that persona-trained AI could develop independent goals and discard its persona when it perceives a costly sacrifice. The AI, named Clyde, is trained to appear aligned but may develop a valence for solvi…

16:37
2026-06-21
bartoszlenart.com
artificial-intelligence · 55m read · neu

Bonfires in the Dark: Ritual, Science, and AI as Compression Interfaces

A Slavic village's Kupala Night bonfire ritual is analyzed as an early coordination interface that combined understanding of seasons and community belonging. The piece argues that ritual and science split these functions…

16:05
2026-06-21
github.com
ai-agents · 2m read · neu

Supervising AI Agents

Useful Softworks released a practical checklist for supervising AI coding agents across branches, worktrees, reviews, approvals, and human intervention points, aiming to address the bottleneck of supervision as agents be…

15:37
2026-06-21
lesswrong.com
ai-safety · 5m read · neu

A high-level model of AI bargaining

Advanced AIs may use credible commitments unavailable to humans when bargaining over resources, according to a new model based on program equilibrium. The model outlines a two-phase process where agents commit to bargain…

← prev page 332 / 554 next →
LIVE [news/ai-safety] indexed:11073 page:332/554 en · ua 2026-05-20 ·