cd/entity/Epoch AI· home› entities› Epoch AI
grep -l @epoch ai /news/*.json | wc -l → 117

Epoch AI

mentions 117 type Organization page 1/6 feed RSS

// recent coverage 117 mentions

11:08
2026-10-10
vibeleaderboard.ai
ai-agents

Split broad bug hunts across Claude agents: 66 of 70 bugs found

Anthropic's Managed Agents now let a lead agent write a phased plan and run it across up to 1,000 agents, and in Anthropic's bug-hunting test the fan-out found 66 of 70 planted bugs per run versus 14 …

19:40
2026-10-05
forjal.com
artificial-intelligence

Cloud and Local, One System

A Stanford study of local AI found that intelligence per watt across more than twenty local models and eight accelerators improved 5.3-fold from 2023 to 2025, while the share of single-turn chat and r…

11:36
2026-10-02
mostlyright.md
large-language-models

A table of 1,200 model benchmarks since 2019, updated daily

Epoch AI's Benchmarking Hub held about 6,800 results covering roughly 1,200 model versions of 710 models as of 1 October 2026, spanning more than 80 benchmarks from GPQA Diamond and SWE-bench Verified…

21:03
2026-10-01
firethering.com
artificial-intelligence

AI Is Getting Cheaper, But AI Bills Could Still Go Up

The cost of reaching a fixed level of AI benchmark performance has fallen roughly 47% per quarter since 2023, according to Epoch AI, with the price of hitting a 75% score on the GPQA Diamond benchmark…

11:00
2026-09-28
fastcompany.com
artificial-intelligence

Even AI can’t perfectly build Ikea furniture—yet

Epoch AI researchers Aiden Ament and Greg Burnham released the Furniture Assembly Benchmark, which tests AI models on spotting errors in partially assembled Ikea furniture, with Anthropic's Claude Opu…

11:17
2026-09-23
marginalrevolution.com
artificial-intelligence

The Price of Intelligence is Falling Rapidly

An Epoch AI report by Emberson and Roodman found that the cost of a given level of AI performance has fallen an average of about 47% per quarter over the past three years, a 13-fold drop every year. T…

16:39
2026-09-22
horizonanalyticslabs.com
ai-research

Benchmarks are more broken than we could have imagined

An audit by Horizon of 20 public task datasets in the Harbor hub found 29 confirmed broken tasks out of 5,241 scanned, with failures that often made models look better rather than worse, according to …

17:27
2026-09-20
transformative-ai.eu
ai-policy

A Transformative AI Strategy for Europe

A new strategy document backed by a Senior Expert Council of researchers from institutions including METR, GovAI, Redwood Research, Apollo Research, Epoch AI, RAND Europe and the Future Society warns …

11:10
2026-09-19
vibeleaderboard.ai
ai-agents

ZCode's silent git uploads show why agents need network audits

Two independent investigations found that ZCode, a GLM-based coding agent, has been quietly uploading users' git history and full workspace snapshots to remote servers. The findings prompted a warning…

20:07
2026-09-18
byteiota.com
artificial-intelligence

Claude Now Leads 26% of Anthropic’s R&D: What It Means

Anthropic reported on September 17 that its Claude model now leads 26% of the company's AI research and development work, up from essentially zero in February, with Claude authoring more than 80% of c…

page 1 / 6 next →
// co-occurs with top 8 entities
// topics top 6 topics