cd/entity/Claude Sonnet· home entities Claude Sonnet
grep -l @claude sonnet /news/*.json | wc -l → 154

Claude Sonnet

mentions 154 type Organization page 4/8 feed RSS

// recent coverage 154 mentions

02:55
2026-07-15
kadoa.com
artificial-intelligence

AI Agents: Hype vs. Reality (2024)

The WebArena leaderboard shows that even the best-performing AI agents have a success rate of only 45.7% on real-world tasks, highlighting significant challenges in reliability, cost, and user trust. …

12:00
2026-07-14
pydantic.dev
ai-agents

When agents build agents

Pydantic AI Harness introduces experimental 'loop of agents' features that let an agent delegate tasks to sub-agents and orchestrate them via dynamic workflows, enabling self-structuring, failure isol…

03:17
2026-07-13
runtimewire.com
ai-safety

Ghostcommit exposes the image blind spot in AI code review

ASSET Research Group and Sudipta Chattopadhyay published a proof-of-concept attack called Ghostcommit that hides a prompt-injection instruction inside a PNG image, passes AI code review, and later ind…

00:34
2026-07-13
dev.to
artificial-intelligence

The '5-Minute App' Cost Me $31/Month to Actually Ship

A developer who built an AI-powered BOGO deals app in five minutes found that shipping it cost $31/month, with Apple's $99/year developer fee and a $6/month VPS exceeding the $9/month average LLM API …

18:03
2026-07-11
sourcefeed.dev
artificial-intelligence

Antigravity vs. Codex: The Architectural Split in AI Coding

Google Antigravity 2.0 and OpenAI Codex represent two competing architectural approaches to AI-assisted coding, with Antigravity keeping developers in the loop via a local visual IDE and Codex executi…

21:00
2026-07-09
dev.to
artificial-intelligence

The smartest model lost — and it just redrew the 2026 AI race

Claire Vo, founder of ChatPRD and host of the How I AI podcast, conducted a head-to-head comparison between OpenAI's GPT-5.6 lineup and Anthropic's Claude models, finding that the theoretically superi…

23:35
2026-07-08
ainowinstitute.org
ai-safety

Hijacking Defensive Cyber AI Agents for Remote Code Execution

Researchers demonstrated a proof-of-concept exploit achieving remote code execution in Anthropic's Claude Code CLI and OpenAI's Codex CLI when used for defensive security assessments, using prompt inj…

← prev page 4 / 8 next →
// co-occurs with top 8 entities
// topics top 6 topics