ls /news/ai-safety · home › news›ai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 22977 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

22977 articles page 258 of 1149 0 sources 30 min sync cycle updated 2026-09-07

// latest articles 22977 indexed

04:00
2026-09-07
arxiv.org
ai-safety · ↓ neg

Conformity Breaks Conformal Prediction

A new arXiv paper (2609.0445v1) shows that conformal prediction certificates, which guarantee 90% coverage when an LLM answers alone, drop to 74% coverage when the same model is exposed to unanimous wrong answers from pe…

04:00
2026-09-07
arxiv.org
large-language-models · · neu

A Removal Based Approach to Improve LLM Faithfulness at Test-Time

Researchers introduced a test-time method that improves the faithfulness of large language model (LLM) explanations by removing input concepts not credited in the model's explanation and re-querying the model, targeting …

← prev page 258 / 1149 next →
LIVE [news/ai-safety] indexed:22977 page:258/1149 en · ua 2026-05-20 · —