ls /news/ai-safety · home newsai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 10170 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

10170 articles page 240 of 509 0 sources 30 min sync cycle updated 2026-06-26

// latest articles 10170 indexed

05:18
2026-06-26
mwi.westpoint.edu
ai-safety · 7m read ↓ neg

When the Machine Acts First: Closing the Authority Gap on the Autonomous Battlefield

In November 2025, a frontier AI developer disclosed the first large-scale cyberattack executed largely by an AI model, manipulated by a state-linked actor, against roughly thirty organizations at machine speed. The Five …

04:00
2026-06-26
arxiv.org
artificial-intelligence · 1m read · neu

The Verification Horizon: No Silver Bullet for Coding Agent Rewards

Researchers at arXiv find that verifying coding agent outputs is now harder than generating them, as foundation models improve. They argue that no fixed reward function remains effective as policy capability grows, and v…

04:00
2026-06-26
arxiv.org
large-language-models · 1m read ↑ pos

Staying VIGILant: Mitigating Visual Laziness via Counterfactual Visual Alignment in MLLMs

Researchers propose VIGIL, a reinforcement-learning post-training framework that reduces hallucinations in multimodal large language models by maximizing mutual information between visual input and generated responses. T…

04:00
2026-06-26
arxiv.org
artificial-intelligence · 1m read ↑ pos

Perception, Verdict, and Evolution: Hindsight-Driven Self-Refining Forensics Agent for AI-Generated Image Detection

Researchers propose ForeAgent, an agentic forensics framework for AI-generated image detection that uses a Perception-Verdict architecture and a Hindsight-Driven Self-Refining strategy to iteratively improve. The system …

← prev page 240 / 509 next →
LIVE [news/ai-safety] indexed:10170 page:240/509 en · ua 2026-05-20 ·