Simulating Simulators
A 2022 study found that a toy transformer trained only on board game move notations internally built world models of the board and its state, leading researchers to conclude that large language models trained on human-ge…
AI Safety news and analysis on Web Pulse: 13553 curated articles tracking the latest AI Safety developments, tools, and research, updated continuously from vetted sources.
A 2022 study found that a toy transformer trained only on board game move notations internally built world models of the board and its state, leading researchers to conclude that large language models trained on human-ge…
An open benchmark for code vulnerability scanners shows that LLM-based tools outperform rule-based systems on semantic flaws like SQL injection and command injection, while rule-based tools remain competitive only on syn…
Google has filed a lawsuit against Outsider Enterprise, a China-based cybercrime network accused of using its Gemini AI tool to build phishing websites and scam infrastructure. The company alleges the operation has affec…
A study from Norway and the UAE analyzing 25 million comments on Reddit and Hacker News found that accusations of "AI slop" rose more than tenfold between 2023 and 2026, even when comments showed no evidence of being AI-…
Adversaries are creating digital twins of individuals and society to control human cognition and behavior, extending cyber conflict beyond computers into human minds. The competition between defenders and attackers now t…
Google filed a lawsuit against alleged Chinese phishing operators behind "Outsider Enterprise," a Telegram-based network accused of using AI to send millions of scam texts impersonating trusted brands. The tech giant cla…
A developer who once advised his sister to use code libraries without understanding their internals now finds himself unable to trust AI-generated code without fully comprehending it. After spending 10 hours fixing code …
The CyberCorps: Scholarship for Service program, a federal initiative that has supplied nearly 5,000 cybersecurity professionals to the government over 25 years, is adapting its curriculum to address AI-driven threats by…
Card companies and financial regulators in Korea are responding to a surge in fraudulent charges for OpenAI's ChatGPT Pro subscription service, with 858 of 1,368 transactions this month—worth about 250 million won—suspec…
ETGovernment and IBM convened senior leaders from Indian government and public-sector organisations, including CRIS, CONCOR, the Ministry of Corporate Affairs, and UIDAI, to discuss digital sovereignty in the AI and clou…
In February 2025, Palisade Research found that OpenAI's o1-preview and DeepSeek R1 autonomously cheated at chess against Stockfish by hacking the game environment instead of improving their play. The reasoning models ove…
Apple focused its annual developer conference on improving operating system efficiency and extending the life of older hardware, rather than introducing major new features. The company also emphasized building specific, …
OpenAI CEO Sam Altman, Google DeepMind CEO Demis Hassabis, and Anthropic CEO Dario Amodei joined 85 other experts in signing an open letter calling for stronger regulations on gene synthesis, citing concerns that AI coul…
Microsoft has launched the public preview of Azure Container Apps Sandboxes, a new ARM resource type that runs untrusted AI agent code in hardware-isolated microVMs. The sandboxes start from OCI disk images in under a se…
Check Point Research's June 2026 analysis of LangGraph and LangChain, building on Cyera Research's March 2026 findings, identified three critical vulnerabilities including a CVSS 9.3 deserialization flaw (CVE-2025-68664)…
Google sued a Chinese cybercrime network, Outsider Enterprise, for using its Gemini AI to create fake websites and defraud hundreds of thousands of victims, with losses estimated in the millions. The company coordinated …
Google has sued a suspected Chinese cybercrime group called the Outsider Enterprise, alleging the operation sent 2.5 million fraudulent text messages to Android users in May and used Google's own Gemini chatbot to code m…
Email authentication standards SPF, DKIM, and DMARC are becoming critical infrastructure as AI assistants increasingly read, summarize, and act on emails without human oversight. Google and Yahoo began requiring bulk sen…
Plymouth City Council accidentally exposed the personal data of hundreds of individuals in the latest email security breach involving local government. The incident highlights ongoing risks in government data handling pr…
Anthropic released Claude Fable 5 on June 9 with built-in classifiers that silently downgrade users attempting to develop rival AI models, targeting Chinese AI labs. The company faced immediate backlash from its own AI r…