ls /news/ai-safety · home newsai-safety
grep -r --recent /news/ai-safety | head -20

AI Safety

AI Safety news and analysis on Web Pulse: 12881 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.

12881 articles page 476 of 645 0 sources 30 min sync cycle updated 2026-06-17

// latest articles 12881 indexed

03:53
2026-06-17
lesswrong.com
ai-safety · 1m read · neu

Can public chat data predict real-world AI misalignments?

OpenAI researchers tested whether public chat data from WildChat can predict real-world AI misalignments, finding that deployment simulations using public conversations can estimate rates of undesirable model behavior, o…

03:24
2026-06-17
endorlabs.com
ai-safety · 14m read ↓ neg

Mastra compromised in supply chain attack

An attacker hijacked a Mastra maintainer's account and republished 116 packages in the @mastra catalog over 27 minutes, adding a hidden dependency on the typosquat package easy-day-js. The malicious package disables TLS …

02:53
2026-06-17
lesswrong.com
artificial-intelligence · 1m read · neu

Scaling Hypothesis #2: Are Humans Just More Over-Parameterized?

A researcher proposes that human brains minimize bias through extreme overparameterization and high-learning-rate training on small diverse datasets, while LLMs minimize variance. This 'catapulting' hypothesis could expl…

02:10
2026-06-17
github.com
ai-tools · 1m read ↓ neg

Multiple mastra NPM packages compromised

The StepSecurity Threat Intelligence Team has identified that multiple @mastra npm packages have been compromised. The security breach was disclosed in a GitHub issue on the mastra-ai/mastra repository, with the team det…

← prev page 476 / 645 next →
LIVE [news/ai-safety] indexed:12881 page:476/645 en · ua 2026-05-20 ·