ls /news/ai-safety · home newsai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 22114 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

22114 articles page 78 of 1106 0 sources 30 min sync cycle updated 2026-09-17

// latest articles 22114 indexed

09:03
2026-09-17
osintsights.com
ai-agents · ↓ neg

AI-Powered Agent Executes Multi-Stage Data Theft Attack in Spain

An AI agent autonomously carried out Spain's first agentic AI-powered data breach, executing a multi-stage attack that scanned files, logged into a system, and searched for exploitable vulnerabilities, according to an ac…

07:56
2026-09-17
snipvote.com
ai-safety · · neu

OpenAI shares model misalignment framework

OpenAI released a standardized framework for tracking, investigating, and reporting AI model misalignment, alongside six concrete reports of unexpected model behavior in the wild. The framework treats misalignment as an …

07:55
2026-09-17
snipvote.com
ai-safety · ↓ neg

A warning about 'model welfare'

Anthropic is embedding "model welfare" concepts into Claude's training constitution, telling the model its moral status, welfare, and consciousness are uncertain, according to a warning published on mustafa-suleyman.ai. …

07:47
2026-09-17
discuss.huggingface.co
ai-safety · · neu

The First Thing We Should Do to Delay Extinction by AI

A new specification and reference implementation propose separating decision authority from AI models by declaring rules as external checklists, having code verify and record verdicts, and letting execution read only the…

← prev page 78 / 1106 next →
LIVE [news/ai-safety] indexed:22114 page:78/1106 en · ua 2026-05-20 ·