cd/entity/GPT-5· home entities GPT-5
grep -l @gpt-5 /news/*.json | wc -l → 172

GPT-5

mentions 172 type Organization page 5/9 feed RSS
sameAs · en.wikipedia.org

// recent coverage 172 mentions

10:11
2026-07-14
machinebrief.com
artificial-intelligence

Anthropic's Claude Fable 5: The Reluctant Genius in Biomedical AI

Anthropic's Claude Fable 5, the company's most capable publicly available model, achieves top-tier accuracy on eight biomedical benchmarks but refuses to answer between 8.0% and 99.4% of questions dep…

10:11
2026-07-14
machinebrief.com
artificial-intelligence

UNIBROWSE Sets New Benchmarks in Multimodal Browsing

UNIBROWSE, a new unified data pipeline for multimodal browsing, achieves an average accuracy of 54.4 on five benchmarks, outperforming its predecessor Qwen3.5-35B-A3B by 10.5 points and surpassing clo…

07:01
2026-07-13
dev.to
artificial-intelligence

What Building an AI Detector Taught Me About False Positives

A developer building an AI content detector discovered that high accuracy claims are misleading, as even a 0.1% false positive rate can wrongly accuse thousands of real people. The tool's probabilisti…

11:00
2026-07-11
dev.to
artificial-intelligence

Same Symptoms, Different Care

A Mount Sinai research team led by Mahmud Omar and Eyal Klang found that OpenAI's GPT-5 amplifies sociodemographic biases in clinical decision-making, assigning higher urgency and less advanced testin…

19:17
2026-07-10
machinebrief.com
large-language-models

The Myth of Autonomous Agents: Why LLMs Aren't Quite There Yet

Current large language models like GPT-5 and Claude Opus-4.1 achieve only a 40% success rate on the PROBE benchmark for proactive AI agents, revealing significant limitations in autonomous problem-sol…

13:23
2026-07-10
machinebrief.com
ai-safety

Simulating AI Deployment: A New Measure of Safety

Deployment simulation, using de-identified past conversations to test AI models before release, is emerging as a more realistic safety evaluation method. A study of GPT-5-series deployments showed it …

04:57
2026-07-10
blog.andymasley.com
artificial-intelligence

Toward liberal environmentalism

Andy Masley argues that debates over AI's environmental impact are often rooted in normative disagreements, not factual ones, and that political liberalism offers a better framework than environmental…

21:14
2026-07-09
thedeepview.com
artificial-intelligence

Can Claude's new screentime tool make AI healthier?

Anthropic launched a reflection dashboard for its Claude AI assistant, allowing users to track usage patterns, set quiet hours, and receive prompts to reflect on their AI reliance. The feature, availa…

12:07
2026-07-07
sourcefeed.dev
artificial-intelligence

The Silent Traps of OpenAI's Assistants API Migration

OpenAI will shut down the Assistants API on August 26, 2026, forcing migration to the Responses API. The architectural shift from stateful to stateless design introduces silent regressions that can de…

← prev page 5 / 9 next →
// co-occurs with top 8 entities
// topics top 6 topics