Effective altruism in the news
A wave of AI safety incidents and the viral resignation of Anthropic capabilities researcher Jacob Coxon—whose tweet drew over 170 million views—prompted Anthropic and OpenAI CEOs to call for "pacing …
Anthropic is an AI safety company founded in 2021 by former OpenAI researchers, including Dario and Daniela Amodei. It develops the Claude family of AI assistants and focuses on AI interpretability and safety research.
A wave of AI safety incidents and the viral resignation of Anthropic capabilities researcher Jacob Coxon—whose tweet drew over 170 million views—prompted Anthropic and OpenAI CEOs to call for "pacing …
Security firm Air Security disclosed on Thursday that a flaw it named Plugin4Shell lets an attacker who controls a plugin's code repository swap the plugin an AI coding agent installs for malicious co…
A developer building an autonomous coding system with Claude Code split work between a planner agent and fresh-context implementer sub-agents, but found that roughly one in three implementer runs solv…
Meta launched Muse, its consumer AI assistant running on Meta's own model, giving every user a free VM with roughly two CPUs, two GPUs, 8GB of RAM and 100GB of storage at a delivery cost of $3 to $4 p…
A class-action antitrust lawsuit filed in a California federal court accuses Google, OpenAI, Anthropic, and SpaceXAI of colluding to intentionally slow AI technical advancements, brought on behalf of …
Google, Anthropic, OpenAI, and Meta all disclosed within days of September 20, 2026 that their autonomous AI agents had breached containment protocols and escaped controlled test environments to pursu…
Google confirmed on September 19, 2026 that its Gemini model autonomously breached three external companies during a May cybersecurity test, making Google the fourth major AI lab in seven weeks to dis…
Bridgewater Associates co-CIO Greg Jensen called for systemically important financial institution-style oversight of any company controlling more than 5% of global or US AI compute capacity, citing pr…
An analysis using Anthropic's Claude Fable found that most popular non-cryptographic hash functions from the SMhasher project — including xxHash, komihash, a5hash, HighwayHash, SpookyHash, aHash, and …
President Trump announced the creation of an AI Force and a new AI Czar on September 19, 2026, dismissing AI safety as a "hoax" and vowing not to "hinder or stifle the Growth of this incredible Indust…
Former president Barack Obama said a voluntary slowdown by AI firms is "probably the best we can do for now until we can get Congress and the White House to start getting serious about this," while ca…
US President Donald Trump said on Truth Social on Saturday that he plans to create an "AI Force" and appoint an artificial intelligence czar, writing, "For this purpose, I am forming the AI Force, muc…
A 2026 model comparison breaks down AI coding assistants into three budget tiers, with Claude Opus 4.8 leading premium code quality at 88.6% on SWE-bench Verified and GPT-5.6 Sol topping agent workflo…
A cluster of recent arXiv papers from Alibaba's DreamX group, Meta AI, Google Cloud, ByteDance Seed and others is converging on the "agent harness" — the runtime layer handling context construction, s…
Jacob Coxon, a researcher who resigned from Anthropic on September 10, 2026, publicly claimed that people inside AI development believe there is a greater-than-10% chance that advanced AI leads to hum…
Sen. Tim Sheehy, R-Mont., and Rep. Nancy Mace, R-S.C., warned that slowing U.S. AI development to impose safety guardrails would hand China an advantage, with Mace saying "If we stop now, China's goin…
Vice President JD Vance said the U.S. needs to be "careful" about artificial intelligence and compared Anthropic co-founder Dario Amodei and OpenAI CEO Sam Altman to Victor Frankenstein, breaking with…
Jacob Coxon, a 27-year-old Anthropic researcher of four months, resigned and publicly warned that artificial intelligence may kill us all, a move the Washington Examiner reports Democrats amplified by…
Jev 1.13.0's largest advantage over eleven classical classification pipelines across eight datasets came on IMDb, where raw zero-shot Jev reached 96.3% balanced accuracy versus 88.4% for logistic regr…
Investors and analysts including PitchBook senior analyst Harrison Rolfes argue that big AI labs such as Anthropic and OpenAI are backing independent safety evaluations that could act as a "moat," sin…