ls /news/ai-safety · home › news›ai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 23252 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

23252 articles page 302 of 1163 0 sources 30 min sync cycle updated 2026-09-02

// latest articles 23252 indexed

03:20
2026-09-02
cleantechnica.com
autonomous-vehicles · · neu

The Zoox Safety Case Framework

Zoox, an Amazon-owned autonomous vehicle company, published its Safety Case Framework detailing a systems safety process and quantitative risk assessment to ensure its robotaxis are significantly safer than human drivers…

03:18
2026-09-02
dev.to
artificial-intelligence · · neu

The AI Had an Authoritative Source. It Was Still Wrong.

A developer building Eterna Clarity discovered that an AI can cite a real, authoritative source and still make a wrong decision when the source supports a different candidate than the one selected. The fix moved from mod…

03:07
2026-09-02
arxiv.org
artificial-intelligence · ↓ neg

One in three AI scribe notes carries a verified clinical error

A preprint audit of three commercial AI scribes on 142 consultations found that 31.3% of 565 notes carried a verified clinical error, with failures concentrated in allergy and medication information, invented patient ide…

01:30
2026-09-02
aiflash.com
ai-safety · · neu

The Safeguard Worked. Is the LLM System Safer?

A new evaluation framework argues that refusal rates and attack success rates do not measure how much harmful assistance a deployed LLM service still provides, and proposes a complementary metric based on the expected ut…

01:09
2026-09-02
promptcube3.com
ai-safety · · neu

How we actually approach AI alignment and security

A technical article outlines a multi-layered approach to AI alignment and security, emphasizing the need to combat reward hacking in RLHF and move toward mechanistic interpretability for deterministic safety. It details …

← prev page 302 / 1163 next →
LIVE [news/ai-safety] indexed:23252 page:302/1163 en · ua 2026-05-20 · —