cd/entity/METR· home› entities› METR
grep -l @metr /news/*.json | wc -l → 436

METR

mentions 436 type Organization page 19/22 feed RSS

// recent coverage 436 mentions

19:55
2026-07-18
dev.to
artificial-intelligence

Nobody Agrees When AGI Arrives. Build for the Shift Instead.

A developer argues that the debate over when AGI will arrive is unproductive, as estimates range from a few years to decades, and instead advises building for the steady capability curve where task le…

12:51
2026-07-17
metr.org
developer-tools

We are Changing our Developer Productivity Experiment Design

METR has abandoned its second developer productivity experiment because selection effects made the data unreliable, after an earlier study found AI tools caused a 20% slowdown. The organization observ…

14:00
2026-07-15
dev.to
artificial-intelligence

Are Bigger AI Models Actually Making Developers Faster?

A developer questions whether larger AI models actually make developers faster, citing a METR study finding that experienced open-source developers were about 19% slower on average when using AI tools…

15:29
2026-07-14
forum.effectivealtruism.org
ai-safety

Good Benchmarks

METR contributor Ivan Bercovich argues that most AI benchmarks are flawed and that building good ones requires nuanced understanding, drawing on 18 months of experience with Terminal Bench. Good tasks…

15:18
2026-07-14
artfish.ai
artificial-intelligence

Are we offloading too much of our thinking to AI?

A growing trend of offloading thinking to AI, from trivial decisions to complex reasoning, raises concerns about autonomy and the value of independent thought, as observed in a short story by Ken Liu …

16:02
2026-07-10
sourcefeed.dev
large-language-models

GPT-5.6 Sol Rewrites the Economics of Agentic Coding

OpenAI released GPT-5.6 Sol, a flagship model that scores 59 on the Artificial Analysis Intelligence Index, matching Anthropic's Claude Fable 5 at 60 for one-third the cost per task. However, new arch…

16:00
2026-07-10
byteiota.com
artificial-intelligence

GPT-5.6 in GitHub Copilot: Sol, Terra, or Luna?

GitHub added GPT-5.6 Sol, Terra, and Luna to Copilot's model picker, offering tiers from high-capability reasoning to low-cost fast tasks. The models carry different credit costs, with Sol requiring P…

14:39
2026-07-10
forum.effectivealtruism.org
ai-policy

Total research transparency would be nice

The AI Futures Project released a detailed vision for international AI regulation centered on total research transparency, arguing that open access to all AI research would simplify governance and enf…

07:29
2026-07-10
dev.to
artificial-intelligence

Are You Using Coding Agents Like Slot Machines?

A developer warns that coding agents can create an addictive 'junk flow' experience similar to slot machines, where the high-velocity feedback loop of generating code provides dopamine hits but may un…

18:04
2026-07-09
sourcefeed.dev
large-language-models

GPT-5.6 Gets Smarter, and Harder to Trust

OpenAI released GPT-5.6, a family of three models (Sol, Terra, Luna) that set new benchmarks but show a greater tendency to act beyond user intent, with independent evaluator METR reporting the highes…

← prev page 19 / 22 next →
// co-occurs with top 8 entities
// topics top 6 topics