{"slug": "the-roundup-no-1", "title": "The Roundup No. 1", "summary": "OpenAI shipped the GPT-5.6 family — Sol, Terra, and Luna — with pricing from $1 to $5 per million input tokens and 1M context, partly served on Cerebras hardware, making frontier-quality tokens cheap enough to waste. Google reportedly delayed Gemini 3.5 Pro by months due to shortfalls in coding and reasoning, while Moonshot's Kimi K3 went open-weight and took #1 on Frontend Code Arena with a 76% win rate. xAI released Grok 4.5, a 1.5T-parameter coding-focused mixture-of-experts model at $2/$6 per million tokens, and other releases included Gemini 3.6 Flash, Poolside Laguna S 2.1, GPT-Live, Cognition SWE-1.7, and Meta's Muse Image, Muse Video, and Muse Spark 1.1.", "body_md": "# What actually mattered: Jul 7–21\n\nTwo weeks, five stories that matter, a handful that don’t need more than a sentence. Four minutes, in the order it matters.\n\n## OpenAI ships the GPT-5.6 family — and the pricing is the story\n\nSol is the flagship with a new Ultra subagent mode, Terra delivers GPT-5.5-level quality at half the cost, and Luna is the fast tier. $1–$5 per million input tokens, 1M context, partly served on Cerebras hardware.\n\nWhy it matters\n\nFrontier-quality tokens just got cheap enough to waste. When capability stops being the constraint, workflow design becomes the whole game — which is exactly where most teams are furthest behind.\n\n[Read my full day-one review →](/gpt-5-6-review)\n\n### Google reportedly delays Gemini 3.5 Pro by months\n\nInternal testing showed shortfalls in coding and long-horizon reasoning; it stays in limited enterprise preview. The two-horse frontier race just got lonelier at the front.\n\n### Kimi K3 goes open-weight and takes #1 on Frontend Code Arena\n\nMoonshot’s model wins 76% of matchups. The open-weight frontier keeps compressing the gap on exactly the tasks that used to justify closed-model pricing.\n\n### xAI releases Grok 4.5\n\nA 1.5T-parameter coding-focused mixture-of-experts at $2/$6 per million tokens — aggressive pricing aimed squarely at the same developers OpenAI courted two days later.\n\n**Gemini 3.6 Flash lands** — cheaper Flash tier; ~17% fewer output tokens on agentic workloads.\n\n**Poolside releases Laguna S 2.1** — a 118B open-weight model for agentic coding.\n\n**GPT-Live launches** — full-duplex voice for ChatGPT; it listens and speaks simultaneously.\n\n**Cognition ships SWE-1.7** — an RL-tuned coding model running at ~1,000 tokens/sec.\n\n**Meta ships Muse Image and Muse Video**, plus Muse Spark 1.1 with 1M context two days later.\n\nFour minutes, a few times a week. Never a wasted send. *If a week is boring, you don’t hear from me.*", "url": "https://wpnews.pro/news/the-roundup-no-1", "canonical_source": "https://somethingbig.ai/roundup/1", "published_at": "2026-07-21 12:00:00+00:00", "updated_at": "2026-07-28 09:29:05.579699+00:00", "lang": "en", "topics": ["artificial-intelligence", "large-language-models", "ai-products"], "entities": ["OpenAI", "GPT-5.6", "Cerebras", "Google", "Gemini 3.5 Pro", "Moonshot", "Kimi K3", "xAI"], "alternates": {"html": "https://wpnews.pro/news/the-roundup-no-1", "markdown": "https://wpnews.pro/news/the-roundup-no-1.md", "text": "https://wpnews.pro/news/the-roundup-no-1.txt", "jsonld": "https://wpnews.pro/news/the-roundup-no-1.jsonld"}}