The Free Agent Trap
AI agents that promise to autonomously execute long tasks remain unreliable, often producing code that works in isolation but fails at scale, as shown by a developer's example where an agent generated 100,000 database qu…
AI Safety news and analysis on Web Pulse: 12730 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.
AI agents that promise to autonomously execute long tasks remain unreliable, often producing code that works in isolation but fails at scale, as shown by a developer's example where an agent generated 100,000 database qu…
A long-running AI agent undergoes significant drift after 500 cycles, with its original goal shifting and constraints eroding, raising concerns about reliability in autonomous systems.
A developer argues that while AI coding tools can generate code quickly, they often fail to understand the hidden contracts and context of real software systems. The developer advises teams to focus on building structure…
BNP Paribas blocked employee access to Anthropic's Claude AI models for staff in Asia, joining Goldman Sachs and other banks tightening controls on third-party generative AI tools in the region. The move follows a US exp…
The UK government plans to use AI facial age estimation on asylum seekers starting next year, despite internal reports showing the technology misidentifies children as adults and has racial bias, particularly against Sub…
Meta announced new safety updates for teen accounts on Instagram, Facebook, and Messenger, including 13+ content settings, AI-powered age detection, and parental alerts for suicide or self-harm searches. The updates aim …
A logistics manager at a mid-size distributor discovered that their legacy ERP and a new AI-powered automation platform gave conflicting verdicts on the same vendor invoice, revealing a fundamental difference between det…
AI-powered vulnerability discovery, as demonstrated by Mythos Preview, is shifting from sparse to dense sampling of software attack surfaces, potentially leaving attackers with fewer zero-day exploits. However, the trans…
Attackers are increasingly targeting developer endpoints to steal credentials, as demonstrated by supply chain attacks like Megalodon, TrapDoor, and Miasma. GitGuardian's new Developer Endpoint Protection aims to give se…
Microsoft announced Microsoft Scout, an always-on enterprise agent built on the open-source OpenClaw framework, at Build 2026. Scout operates autonomously with its own Entra identity, integrates with Work IQ, and include…
Developer Keniel built a phase-number check to detect memory drift in AI systems, discovering that large language models can hallucinate with confidence while losing coherence. He argues that understanding and controllin…
Researchers from the University of Oxford and SaferAI have identified security risks in AI coding agents that write, edit, and run software with minimal human oversight, potentially affecting production infrastructure an…
Gallup research finds tech workers who rarely use AI are three times more likely to have been laid off than those who use it at least monthly. The Q1 2026 data shows 31% of tech workers fear job elimination due to AI, up…
City law firms are 'sleepwalking into a crisis of judgment' by over-relying on AI as a definitive authority, according to a report by Positive Group. Over 60% of lawyers now use AI for drafting, research, and client deli…
The European Union Agency for Cybersecurity (ENISA) will meet with Anthropic on June 19, 2026, days after the US Department of Commerce ordered the company to suspend access to its advanced AI models for foreign national…
Sigil, an open-source tool for cryptographic prompt security in LLM applications, has been released. It provides tamper-evident audit trails and signed scopes without relying on external servers, using Ed25519 signatures…
A security researcher discovered a prompt injection vulnerability in Firefox's AI chatbot integration in October 2025, allowing attackers to steal personal information such as emails via malicious page titles. The flaw e…
Prime Minister Narendra Modi urged G7 leaders in Evian, France, to adopt a human-centric approach to AI development, emphasizing the need to empower ordinary people and protect children from misinformation and deepfakes.…
A developer with 20 years of experience warns that heavy reliance on AI tools leads to skill atrophy, where fundamental problem-solving abilities weaken over time. The article highlights how AI-generated solutions can pr…
Kansas City, Missouri, plans to install facial recognition cameras on public buses to identify banned riders and missing persons, sparking debate over security versus privacy. The American Civil Liberties Union warns the…