cd/entity/GPT-5.6 Luna· home› entities› GPT-5.6 Luna
grep -l @gpt-5.6 luna /news/*.json | wc -l → 196

GPT-5.6 Luna

mentions 196 type Person page 5/10 feed RSS

// recent coverage 196 mentions

07:24
2026-08-18
byteiota.com
ai-infrastructure

AI API Price War: What Developers Must Act on in August

DeepSeek quadrupled output prices for V4-Pro and V4-Flash on August 16, 2026, raising V4-Pro peak output from $0.87 to $3.96 per million tokens, a 355% spike, while Anthropic permanently locked Claude…

05:00
2026-08-18
theaq.blog
artificial-intelligence

Evaluating Hy3 on Hack The Box Challenges

Tencent's Hy3 scored 34.8% on the HTB-Challenger Benchmark, the third-lowest among all tested models, and got stuck on 9 of 16 Hack The Box challenges, generating a median of 101,509 output tokens per…

01:41
2026-08-18
shukla.io
ai-research

Who benchmarks the benchmark?

A new audit of the EnterpriseOps Gym benchmark found that fixing environment issues in the 'Teams' domain raised GPT-5.6 Luna's score from 26.2% to 100% on 61 tasks, revealing that many agent failures…

06:40
2026-08-17
dev.to
large-language-models

The Model Knew the Bid Was True. Then It Challenged Anyway.

In Kai, a Liar's Dice game, an engineer found that large language models sometimes challenge a bid they know is true, losing the round. The issue was traced to the action schema, where the 'challenge'…

18:36
2026-08-16
cryptobriefing.com
artificial-intelligence

US labs cut AI inference costs nearly 25% amid price war

Average prices for AI inference from leading US labs dropped nearly 25% between mid-July and mid-August, according to Silicon Data analysis reported by the Financial Times. OpenAI cut prices on GPT-5.…

00:00
2026-08-16
martinalderson.com
artificial-intelligence

How I think about reducing AI costs

AI inference costs are becoming a major problem for many companies, with AI spend per employee per month rising sharply, according to the Ramp AI Index via a16z. To reduce costs, companies should audi…

14:30
2026-08-14
digitalapplied.com
artificial-intelligence

LLM Batch APIs: The Half-Price Lane Nobody Budgets

Google, OpenAI, and Anthropic all charge exactly half their synchronous rate for batch API work, yet most cost forecasts price every token at the sync list rate, creating a 2× error on workloads that …

19:52
2026-08-13
thedeepview.com
artificial-intelligence

Google wagers Gemini can win on economics

Google announced Gemini 3.7 Flash, its most intelligent workhorse model for coding and agents, priced at $0.75 per million input tokens and $3.75 per million output tokens, matching its predecessor Ge…

19:03
2026-08-13
dev.to
large-language-models

Local vs Hosted LLMs: The Decision Framework

A developer has published a decision framework for choosing between local and hosted large language models, covering cost, privacy, latency, and control. The guide includes tools like a break-even cal…

← prev page 5 / 10 next →
// co-occurs with top 8 entities
// topics top 6 topics