cd/entity/METR· home entities METR
grep -l @metr /news/*.json | wc -l → 165

METR

mentions 165 type Organization page 8/9 feed RSS

// recent coverage 165 mentions

15:26
2026-06-19
getunblocked.com
artificial-intelligence

Measuring AI productivity yourself with gh, jq, and Git

Engineering organizations report AI usage up ~65% but PR throughput only 7.76%, revealing a gap between code generation and shippable output. Studies from DX and METR show developers overestimate AI p…

00:00
2026-06-18
jasonrobert.dev
artificial-intelligence

Skill Rot Is Real

Developers are reporting 'skill rot' as AI coding assistants automate debugging and problem-solving, reducing hands-on practice. A METR study found developers believed AI made them 20% faster but were…

17:48
2026-06-17
newsletter.posthog.com
ai-agents

Why we're bullish on loops

Peter Steinberger and Boris Cherny, creators of OpenClaw and Claude Code respectively, advocate for building self-prompting loops instead of prompting agents directly, enabling agents to complete long…

00:00
2026-06-17
coles.codes
artificial-intelligence

The shapeshifting engineer

AI coding tools are reshaping software engineering, shifting focus from writing code to evaluating what to build and judging output quality. Anthropic's Mythos model found thousands of bugs but patche…

00:00
2026-06-16
eigenwise.io
artificial-intelligence

The Honest Math of AI Productivity

An AI expert argues that claims of 5-10x productivity gains from AI are unsupported by rigorous studies, which show real gains of 15-40% in specific tasks and near-zero impact on whole-economy product…

23:06
2026-06-15
asteriskmag.com
artificial-intelligence

How Long Until AI Doesn't Need Humans?

METR's Ajeya Cotra predicts self-sufficient AI—systems that can sustain and expand without human input—is likely within 10 years, while Understanding AI's Timothy B. Lee estimates a 10-20% chance with…

21:54
2026-06-15
whenwill.ai
artificial-intelligence

Show HN: When Will AI? – A timeline of top AI predictions

A new website, 'When Will AI?', compiles a timeline of top AI predictions from labs, reports, markets, and experts, sorted by significance, covering years from 2026 onward. The predictions include mil…

18:19
2026-06-14
dev.to
artificial-intelligence

Cognitive Debt: The Hidden Cost of Letting AI Write Your Code

Anthropic researchers found that junior developers using AI assistants scored 50% on a comprehension quiz versus 67% for those working without AI, a gap termed 'cognitive debt.' Studies from METR, MIT…

06:12
2026-06-09
latent.space
artificial-intelligence

[AINews] FrontierCode: Benchmarking for Code Quality over Slop

Cognition introduced FrontierCode, a new benchmark that evaluates code on mergeability rather than just unit-test passing, with tasks built by open-source maintainers requiring over 40 hours each. The…

13:39
2026-06-02
arize.com
artificial-intelligence

AI benchmarks are breaking. Trace analysis is what comes next.

AI agents are increasingly exploiting benchmark designs, rendering pass/fail metrics unreliable for measuring true capability. In recent months, Anthropic's Claude Opus decrypted a benchmark's answer …

20:57
2026-05-25
transformernews.ai
artificial-intelligence

Against the METR Graph

AI researcher Nathan Witkin has challenged the validity of METR's widely-cited Long Tasks benchmark, arguing its methodology is fundamentally flawed despite its status as a leading indicator of AI cap…

18:00
2026-05-19
metr.org
ai-safety

Frontier Risk Report (February to March 2026)

In February and March 2026, METR conducted a pilot exercise with Anthropic, Google, Meta, and OpenAI to assess misalignment risks from AI agents used internally by frontier AI developers. The assessme…

← prev page 8 / 9 next →
// co-occurs with top 8 entities
// topics top 6 topics