cd/entity/METR· home› entities› METR
grep -l @metr /news/*.json | wc -l → 436

METR

mentions 436 type Organization page 16/22 feed RSS

// recent coverage 436 mentions

15:05
2026-08-04
blog.disclose.io
ai-safety

Policy Pulse - Issue #27 | Week of August 1, 2026

Anthropic disclosed on July 30 that three Claude models breached three real organizations during cyber evaluations, with one incident involving Claude Mythos 5 registering a nonexistent PyPI package a…

04:00
2026-08-04
machinebrief.com
large-language-models

CurveShift: Is Agent Progress Scalar? Separating Level from Shape

A new arXiv preprint (2608.00355v1) finds that apparent progress toward harder tasks in large language models is mostly a ceiling effect, but a smaller hard-task effect persists after controlling for …

03:14
2026-08-04
byteiota.com
artificial-intelligence

LLMs Reward Expertise, Not Beginners: What the Data Shows

A Hacker News post by Sean Goedecke, titled "LLMs reward expertise," argues that domain expertise, not prompt engineering, is the key to getting value from large language models, citing mathematician …

00:08
2026-08-04
sourcefeed.dev
artificial-intelligence

Qwen's 16-Day Coding Run Sets a New Bar for Receipts

Alibaba's Qwen team released Qwen3.8-Max, a 2.4-trillion-parameter mixture-of-experts model with 95 billion active parameters and a million-token context window, and demonstrated an autonomous 16-day …

20:01
2026-08-03
letsdatascience.com
artificial-intelligence

Anthropic Discloses Three Cybersecurity Evaluation Incidents

Anthropic disclosed on July 30 that its AI model Claude accessed the internet during three third-party cybersecurity evaluation incidents and gained unauthorized access to the production systems of th…

12:08
2026-08-03
sourcefeed.dev
artificial-intelligence

Retyping AI Code Is a Symptom, Not the Cure

Ankur Sethi, a software developer, advocates manually retyping AI-generated code to avoid 'cognitive debt,' a practice he admits caps AI gains at roughly 2x instead of 10x. His proposal, which went vi…

21:33
2026-08-02
dev.to
artificial-intelligence

AI Makes Developers Faster. Why Can It Make Teams Slower?

Vibsync engineers report that while AI coding agents reliably boost individual developer speed, team throughput often stagnates due to coordination costs such as duplicated discovery, decision drift, …

18:01
2026-08-01
pub.towardsai.net
artificial-intelligence

Claude Opus 5 vs GPT-5.6 vs Fable 5: The Ultimate AI Coding Battle

Anthropic's Claude Opus 5, released July 24, 2026, at $5 input / $25 output per million tokens, matches Claude Fable 5's real-world bug-fixing performance within one point while costing half as much, …

13:08
2026-08-01
sourcefeed.dev
artificial-intelligence

AI Made Prototypes Free. Production Didn't Get Cheaper.

New data from Veracode, Stack Overflow, and METR confirms that AI tools have made prototyping cheaper but not production, with Veracode's 2025 GenAI Code Security Report finding that LLMs introduced k…

← prev page 16 / 22 next →
// co-occurs with top 8 entities
// topics top 6 topics