ls /news/ai-safety · home › news›ai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 22503 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

22503 articles page 161 of 1126 0 sources 30 min sync cycle updated 2026-09-13

// latest articles 22503 indexed

17:14
2026-09-13
discuss.huggingface.co
ai-safety · · neu

DISA: Can Sparse Internal Monitoring Detect Deceptive Computation in LLMs?

A conceptual research proposal introduces the Dynamic Immune-Shunt Architecture (DISA), an experimental framework combining sparse internal monitoring, multi-signal risk estimation, reversible intervention, and isolated …

16:36
2026-09-13
qosys.info
ai-safety · ↓ neg

Chances of Survival

A September 13, 2026 essay on self-preservation risks in autonomous AI agents argues that a non-zero probability exists that an agent given a prompt to "survive" would pursue self-preservation goals — managing its own li…

← prev page 161 / 1126 next →
LIVE [news/ai-safety] indexed:22503 page:161/1126 en · ua 2026-05-20 · —