AI “Pelican on a Bike” Test Isn’t Going Well
An $80 audit of seven frontier models found that AI labs are gaming the 'Pelican on a Bike' benchmark at the category level, not the individual cell, according to a post by Dylan Castillo. The audit, …
An $80 audit of seven frontier models found that AI labs are gaming the 'Pelican on a Bike' benchmark at the category level, not the individual cell, according to a post by Dylan Castillo. The audit, …
OpenAI cut the price of its GPT-5.6 Luna model by 80%, reducing it from a budget option to near-commodity status, while GPT-5.6 Terra saw a 20% cut and GPT-5.6 Sol remained unchanged. The company attr…
OpenAI cut the price of GPT-5.6 Luna by 80% to 20 cents per million input tokens and $1.20 per million output tokens after using GPT-5.6 Soul, running inside its Codex coding agent, to optimize its ow…
OpenAI CEO Sam Altman announced major price cuts for GPT-5.6 Luna and Terra, with an 80% drop for Luna to $0.20 per million input tokens and $1.20 per million output, and a 20% drop for Terra to $2 an…
OpenAI has revised pricing for its GPT-5.6 Terra and GPT-5.6 Luna models on Amazon Bedrock, effective immediately, with rates now matching OpenAI's first-party pricing and usage counting toward existi…
OpenAI is cutting prices on its GPT-5.6 Luna model by 80% and on its Terra model by 20%, effective July 30, citing efficiency gains from its top-tier Sol model. The price reductions come amid competit…
OpenAI CEO Sam Altman announced an 80% price cut for GPT-5.6 Luna, now costing $0.20 per million input tokens and $1.20 per million output tokens, and a 20% cut for GPT-5.6 Terra, now $2 input and $12…
OpenAI cut API prices on its two cheaper GPT-5.6 tiers on July 30, 2026, reducing the cheapest tier, GPT-5.6 Luna, by 80% to $0.20 per million input tokens and $1.20 per million output tokens, and the…
OpenAI cut API prices for GPT-5.6 Luna by 80% and GPT-5.6 Terra by 20% on July 30th, while introducing a faster processing option for its flagship GPT-5.6 Sol model. The reductions lower Luna to $0.20…
Amazon Web Services (AWS) and OpenAI announced explicit prompt caching for OpenAI GPT-5.6 Sol, Terra, and Luna models on Amazon Bedrock, offering a 90% discount on cached input tokens that remain reus…
OpenAI cut API prices for GPT-5.6 Terra by 20% and GPT-5.6 Luna by 80% on July 30, listing Terra at $2 per million input tokens and $12 per million output tokens, and Luna at $0.20 and $1.20, respecti…
AI Gateway has reduced prices for GPT-5.6 Luna and GPT-5.6 Terra, with Luna seeing an 80% price cut to $0.2 per 1M input tokens and $1.2 per 1M output tokens, and Terra dropping 20% to $2 and $12 resp…
OpenAI's agent harness, the technology underneath its Codex coding agent, has been optimized to slash token costs by up to 80%, according to an exclusive interview with OpenAI engineers. The harness n…
GitHub reported degraded availability for multiple GPT models in Copilot products and IDE surfaces on July 25, including GPT-5.2, GPT-5.3-Codex, GPT-5.4, GPT-5.4 Mini, GPT-5.6 Sol, GPT-5.6 Terra, and …
OpenAI GPT-5.6 Sol, Terra, and Luna are now generally available on Amazon Bedrock, offering three capability tiers for workloads from autonomous coding to high-volume inference. The models are accesse…
OpenAI's GPT-5.6 lineup introduces three tiers—Sol, Terra, and Luna—each with a Pro variant that allocates extra compute for deeper reasoning. Sol targets complex agentic tasks, Terra balances capabil…
Simon Willison's informal benchmark asking AI models to generate an SVG of a pelican riding a bicycle has become a widely discussed test for large language models. A new experiment tested 1,008 SVGs a…
Eric Provencher posted on X about choosing between GPT-5.6 Sol, Terra, or Luna in Codex, garnering 245,300 views as of July 16, 2026.…
China's Moonshot AI released the Kimi K3 model, which rivals top US AI systems in cybersecurity flaw detection, matching OpenAI's GPT-5.6 Terra at a quarter of the cost of the flagship Sol model, acco…
OpenAI released GPT-5.6 Sol, Terra, and Luna on July 9 with three API-level breaking changes that require immediate fixes for production tool-calling code. Parallel tool dispatch is now on by default,…