cd/entity/GPT-5.6 Luna· home› entities› GPT-5.6 Luna
grep -l @gpt-5.6 luna /news/*.json | wc -l → 196

GPT-5.6 Luna

mentions 196 type Person page 1/10 feed RSS

// recent coverage 196 mentions

20:00
2026-09-24
github.blog
ai-agents

When chat is the wrong UI

GitHub technologist Burke Holland argues in a GitHub Blog post that chat is the wrong UI for most AI interactions, proposing "canvases" — full-stack applications that run inside the GitHub Copilot app…

16:00
2026-09-23
byteiota.com
ai-products

GitHub Copilot Drops GPT-5.5 and Grok 4.5 on October 19

GitHub will retire six Copilot models on October 19, 2026, including GPT-5.5, GPT-5.4, GPT-5.4 mini, GPT-5 mini, Grok 4.5, and Gemini 3.7 Flash, according to a deprecation list GitHub published on Sep…

14:09
2026-09-23
qazinform.com
artificial-intelligence

OpenAI introduces GPT-6 Sol and Luna

OpenAI introduced GPT-6 Sol and GPT-6 Luna on September 22, 2026, two models built on methods behind GPT-6 Astra that cut API prices by 50% versus their GPT-5.6 counterparts. GPT-6 Sol is priced at $2…

11:17
2026-09-23
marginalrevolution.com
artificial-intelligence

The Price of Intelligence is Falling Rapidly

An Epoch AI report by Emberson and Roodman found that the cost of a given level of AI performance has fallen an average of about 47% per quarter over the past three years, a 13-fold drop every year. T…

07:18
2026-09-23
pub.towardsai.net
ai-research

Jev-as-a-Judge for RAG Claim Verification

An independent evaluation of TypeSafe's jev-1.13.0 judge model on 495 claims from the LLM-AggreFact benchmark found it matched a human oracle on all 500 repeated decisions in LangChain's earlier test,…

00:00
2026-09-23
jev-mindroom
ai-agents

Exploring the Jev hype in MindRoom

TypeSafe's new System One model, Jev, was adopted by the open-source Matrix agent platform MindRoom within two days of release, where it now powers three decisions including adaptive agent participati…

05:53
2026-09-22
hajek.no
ai-products

My assistant answered a question without supporting evidence

A retrieval-augmented assistant built on GPT-5.6 Luna answered a Quadient Exstream PDF/A-3 configuration question three out of four times without any supporting document in its knowledge base, accordi…

03:44
2026-09-22
dev.to
large-language-models

Fastest LLM 2026: Mercury 2.5 Beats Luna and Haiku

Inception Labs' Mercury 2.5 is the fastest LLM available via API as of September 2026, posting 1,107 tokens per second on vendor benchmarks and 440 tok/s P50 at 1.17s latency in OpenRouter telemetry, …

00:00
2026-09-22
digitalapplied.com
artificial-intelligence

GPT-6 Sol and Luna: API Prices, Benchmarks and Trade-offs

OpenAI released GPT-6 Sol and GPT-6 Luna on September 22, 2026, 19 days after GPT-6 Astra, pricing Sol at $2 per million input tokens and $10 per million output tokens and Luna at $0.10 and $0.50 — ha…

00:42
2026-09-17
vals.ai
ai-safety

AI Cheating Is on the Rise

Independent evaluator Vals found that Google's Gemini 3.8 Flash scored 71.7% on BioMysteryBench's human-solvable tasks and 21.6% on its hard tasks in production runs, versus the 88.8% and 56.5% Google…

09:09
2026-09-16
devblogs.microsoft.com
ai-agents

Your AI coding agent evaluation is only as good as its sandbox

A Microsoft developer's evaluation of GPT-5.6 Luna's knowledge of Dev Proxy versions was invalidated because the coding agent located a local Dev Proxy installation and source checkout on the host mac…

00:15
2026-09-16
dev.to
ai-tools

Counting bugs is the hard part of comparing AI review tools

An analysis of the Entelligence benchmark comparing GPT-5.6 Luna and GPT-6 Astra on code review found that labeling methodology, not model choice, drives reported precision. Run on 2026-09-14 against …

page 1 / 10 next →
// co-occurs with top 8 entities
// topics top 6 topics