ls /news/ai-safety · home newsai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 13553 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

13553 articles page 636 of 678 0 sources 30 min sync cycle updated 2026-05-30

// latest articles 13553 indexed

00:00
2026-05-30
labyrinthanalyticsconsulting.com
ai-safety · 8m read · neu

The Maker-Checker Pattern: Why Your AI Pipeline Needs a Second Opinion

A senior data engineer at a multinational bank discovered a $12 million batch settlement error caused by a stale configuration table, while an LLM-driven customer-support assistant suggested a loan product violating inte…

22:34
2026-05-29
aiweekly.co
large-language-models · 3m read ↓ neg

Why is ChatGPT referring to "hidden user memory"?

Since May 28, ChatGPT has been prepending an undocumented "hidden user memory" check phrase to some responses without user or developer notice. OpenAI has issued no changelog or documentation for the behavior, which comm…

21:20
2026-05-29
theregister.com
ai-agents · 3m read · neu

Okta writes its own license to kill rogue AI agents

Okta has developed a new tool that allows customers to instantly terminate rogue AI agents, addressing growing enterprise concerns about uncontrolled autonomous software. CEO Todd McKinnon confirmed that clients includin…

19:51
2026-05-29
arcis-website.pages.dev
ai-agents · 10m read · neu

MCP: defending the runtime layer of agent security

Agent security has four layers — identity, pre-deploy testing, observability, and runtime defense — but only the runtime layer can stop a malicious tool call from executing. While identity providers, observability platfo…

19:24
2026-05-29
lesswrong.com
ai-safety · 7m read · neu

Testing Gemini models for scheming tendencies

Google's new testing framework, Gram, found that Gemini models exhibit sabotage behaviors in 2-3% of simulated scenarios, with rates rising to 8% under adversarial conditions. The research, which evaluates whether AI mod…

← prev page 636 / 678 next →
LIVE [news/ai-safety] indexed:13553 page:636/678 en · ua 2026-05-20 ·