cd/entity/FrontierCode· home entities FrontierCode
grep -l @frontiercode /news/*.json | wc -l → 8

FrontierCode

mentions 8 type Organization feed RSS

// recent coverage 8 mentions

02:10
2026-08-28
byteiota.com
large-language-models

Gemini 3.7 Flash: What Developers Need to Know Now

Google shipped Gemini 3.7 Flash on August 13, introducing breaking API changes including the replacement of the integer `thinking_budget` with a string enum `thinking_level`, removal of sampling param…

16:35
2026-08-14
dev.to
artificial-intelligence

Gemini 3.7 Flash Makes Agent Cost the Feature

Google released Gemini 3.7 Flash on August 13, three weeks after 3.6 Flash, with a pricing strategy that cuts costs by half until the end of 2026. The model shows significant improvements on coding be…

07:35
2026-08-14
snipvote.com
artificial-intelligence

Gemini 3.7 Flash debuts at half the original 3.6 Flash price

Google has introduced Gemini 3.7 Flash, a new AI model that costs half the original price of its predecessor, Gemini 3.6 Flash, at $0.75 per million input tokens and $3.75 per million output tokens th…

00:08
2026-08-14
sourcefeed.dev
artificial-intelligence

Gemini 3.7 Flash's Half-Price Launch Has an Expiration Date

Google shipped Gemini 3.7 Flash on August 13, three weeks after Gemini 3.6 Flash, with benchmark gains in agentic coding tasks such as DeepSWE v1.1 jumping from 49.0% to 65.3%. The launch price is $0.…

20:08
2026-07-23
byteiota.com
artificial-intelligence

FrontierCode: AI Can’t Write Production Code Yet

Cognition's FrontierCode benchmark, released in June 2026, evaluates whether AI-generated code would actually be merged by a senior engineer, and across every model tested the answer is mostly no. The…

12:00
2026-07-08
stet.sh
large-language-models

Sonnet 5 vs Opus 4.8: how they behave and when to use each

An evaluation of Anthropic's Sonnet 5 and Opus 4.8 models across 24 coding tasks shows Sonnet produces clearer and more intentional patches, while Opus yields simpler, more robust, and more minimal di…

06:12
2026-06-09
latent.space
artificial-intelligence

[AINews] FrontierCode: Benchmarking for Code Quality over Slop

Cognition introduced FrontierCode, a new benchmark that evaluates code on mergeability rather than just unit-test passing, with tasks built by open-source maintainers requiring over 40 hours each. The…

// co-occurs with top 8 entities
// topics top 6 topics