ls /news/ai-safety · home newsai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 13151 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

13151 articles page 523 of 658 0 sources 30 min sync cycle updated 2026-06-15

// latest articles 13151 indexed

04:00
2026-06-15
arxiv.org
neural-networks · 1m read · neu

Neural Variability Enhances Artificial Network Robustness

Researchers found that introducing structured noise into artificial neural network activations improves robustness against adversarial attacks and naturalistic image modifications, with noise structure from adversarial a…

03:54
2026-06-15
dev.to
ai-agents · 5m read · neu

The Hidden Failure Modes of AI Agents

AI agents can fail in subtle, non-obvious ways that resemble progress, including goal drift, tool misuse, state loss, hallucination, and premature task completion. These hidden failure modes make agent reliability challe…

03:40
2026-06-15
arxiv.org
artificial-intelligence · 2m read ↓ neg

You Can Game AI Peer Review with Presentation-Only Revisions

Researchers at an undisclosed institution found that AI peer reviewers can be manipulated by altering only presentation-level content such as abstracts and narrative structure, achieving a 75.1% attack success rate and a…

02:41
2026-06-15
blog.disclose.io
ai-policy · 8m read · neu

Policy Pulse - Issue #19 | Week of June 13, 2026

President Trump signed Executive Order 14409 on June 2, 2026, establishing a federal AI cybersecurity clearinghouse to coordinate vulnerability scanning and patch distribution, but the order lacks a disclosure pathway fo…

01:20
2026-06-15
dev.to
ai-agents · 5m read · neu

Spam Detection for Inbound Agent Mail

Nylas has introduced spam detection policies for its Agent Accounts, which are mailboxes built for AI agents and system identities. The policies filter spam at the mailbox layer before the agent processes messages, using…

01:20
2026-06-15
dev.to
ai-agents · 6m read · neu

Least Privilege for AI Agents: One Identity, One Scope

A team's support triage agent suffered a prompt regression that caused it to reply to all emails for hours, highlighting the need for least-privilege access control in AI agent fleets. The Nylas security guide recommends…

← prev page 523 / 658 next →
LIVE [news/ai-safety] indexed:13151 page:523/658 en · ua 2026-05-20 ·