ls /news/ai-safety · home › news›ai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 22992 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

22992 articles page 259 of 1150 0 sources 30 min sync cycle updated 2026-09-07

// latest articles 22992 indexed

04:00
2026-09-07
arxiv.org
ai-safety · ↓ neg

Conformity Breaks Conformal Prediction

A new arXiv paper (2609.0445v1) shows that conformal prediction certificates, which guarantee 90% coverage when an LLM answers alone, drop to 74% coverage when the same model is exposed to unanimous wrong answers from pe…

04:00
2026-09-07
arxiv.org
large-language-models · · neu

A Removal Based Approach to Improve LLM Faithfulness at Test-Time

Researchers introduced a test-time method that improves the faithfulness of large language model (LLM) explanations by removing input concepts not credited in the model's explanation and re-querying the model, targeting …

04:00
2026-09-07
arxiv.org
artificial-intelligence · · neu

Corporate Language Model (CLM): Transforming Tacit and Fragmented Enterprise Knowledge into a Sovereign, Auditable, and Executable Corporate Intelligence Layer

A new arXiv paper (2609.04377v1) introduces the Corporate Language Model (CLM), a framework that converts a firm's tacit and fragmented knowledge into an ontology-grounded enterprise intelligence layer with five capabili…

04:00
2026-09-07
machinebrief.com
artificial-intelligence · · neu

Shadow Queries for Private Retrieval in Vector Databases

Researchers propose SHAQ, a defense against embedding inversion attacks that reconstruct text from vector database embeddings, achieving a recovery rate as low as 0.2104 and defending up to 19.50% more tokens than baseli…

← prev page 259 / 1150 next →
LIVE [news/ai-safety] indexed:22992 page:259/1150 en · ua 2026-05-20 · —