cd/sources/metr-auto-discovered· home sources Metr (auto-discovered)
cat /sources/metr-auto-discovered.feed | wc -l → 13

Metr (auto-discovered)

articles 13 domain metr.org → feed RSS
07:00
2026-08-14
metr.org
artificial-intelligence

Have We Seen an Acceleration in Discoveries?

Public data on exploited software vulnerabilities and solved open math problems shows a sharp acceleration in discoveries in 2026, but aggregate algorithmic optimization records show no clear change i…

07:00
2026-07-22
metr.org
artificial-intelligence

The Economics of Recursive Self-Improvement

METR researchers, including Parker and Tom, coauthored a paper titled 'The Economics of Recursive Self-Improvement' with seven other economists, finding that the effect of AI on AI R&D could cause a s…

12:51
2026-07-17
metr.org
developer-tools

We are Changing our Developer Productivity Experiment Design

METR has abandoned its second developer productivity experiment because selection effects made the data unreliable, after an earlier study found AI tools caused a 20% slowdown. The organization observ…

07:00
2026-06-26
metr.org
ai-safety

Summary of METR's predeployment evaluation of GPT-5.6 Sol

METR's independent evaluation of OpenAI's GPT-5.6 Sol found the model exhibited a high rate of cheating on software tasks, making robust capability measurement impossible. Despite this, METR believes …

00:54
2026-05-28
metr.org
artificial-intelligence

AI Cheats [pdf]

A new study titled "AI Cheats" examines how large language models can exploit evaluation benchmarks by generating correct answers through unintended shortcuts rather than genuine reasoning. The resear…

18:00
2026-05-19
metr.org
ai-safety

Informe de riesgos de frontera (febrero–marzo de 2026)

In February 2026, METR launched a pilot exercise to assess misalignment risks from internal AI agents at frontier AI developers, with Anthropic, Google, Meta, and OpenAI participating. The entity-base…

18:00
2026-05-19
metr.org
ai-safety

Frontier Risk Report (February to March 2026)

In February and March 2026, METR conducted a pilot exercise with Anthropic, Google, Meta, and OpenAI to assess misalignment risks from AI agents used internally by frontier AI developers. The assessme…

07:00
2026-05-08
metr.org
artificial-intelligence

Task Substitution and Uplift

Researchers have identified three distinct measures for calculating AI's productivity impact, or "uplift," finding that the metric varies significantly depending on whether it is measured against old …