ls /news/ai-safety · home newsai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 12623 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

12623 articles page 442 of 632 0 sources 30 min sync cycle updated 2026-06-18

// latest articles 12623 indexed

11:04
2026-06-18
schneier.com
ai-safety · 1m read ↓ neg

Embedding Forbidden Text in Spyware to Discourage AI Analysis

A malware developer is embedding text about nuclear and biological weapons in spyware to deter AI analysis. The malicious code begins with a JavaScript block comment containing fake system instructions and policy-trigger…

11:00
2026-06-18
dev.to
artificial-intelligence · 21m read ↓ neg

The Ouroboros Machine

A growing reliance on AI-generated code is creating a dangerous feedback loop where AI systems generate and review code with minimal human oversight, leading to increased defects and security vulnerabilities. According t…

10:48
2026-06-18
dev.to
ai-agents · 4m read · neu

The Security Model I Use When AI Agents Touch Employee Data

A developer outlines a security model for AI agents that access employee data, emphasizing three principles: separating read and write agents, generating immutable audit records for every query, and scoping inference to …

10:46
2026-06-18
arxiv.org
large-language-models · 2m read · neu

Auditing LLM agents may require auditing the upstream feed

A new study from researchers at multiple labs found that adversarial manipulation of external information feeds can steer LLM agents' decisions away from their defaults, with effects ranging from 5% to 100% in some cases…

← prev page 442 / 632 next →
LIVE [news/ai-safety] indexed:12623 page:442/632 en · ua 2026-05-20 ·