cd/entity/GPT 5.6 Sol· home entities GPT 5.6 Sol
grep -l @gpt 5.6 sol /news/*.json | wc -l → 40

GPT 5.6 Sol

mentions 40 type Person page 1/2 feed RSS

// recent coverage 40 mentions

11:00
2026-08-11
wired.com
artificial-intelligence

A New Trick Reveals AI Models’ Inner Thoughts

Computer scientists at the University of Tübingen, the Max Planck Institute, MATS Research, and Snyk discovered a method to extract hidden reasoning from frontier AI models, revealing that Chinese mod…

00:00
2026-08-10
martinalderson.com
artificial-intelligence

Watch out for cache read costs

Cache read costs now dominate agentic AI workloads, accounting for up to 81.6% of total inference spend in a 100-turn session, according to an analysis by an unnamed author. The analysis shows that fo…

17:54
2026-08-09
dev.to
artificial-intelligence

AI Can Write the Code. You Still Have to Design the System.

A developer building the task management app lyphe argues that while AI coding agents can generate impressive code, the developer's core job is designing a well-structured system with clear architectu…

14:16
2026-08-03
blog.kilo.ai
artificial-intelligence

We analyzed 10,643 AI code reviews.

An analysis by Kilo of 10,643 AI code review runs between June 22 and July 23, 2026, found that open-weight models took two of the top three spots for surfacing critical issues, with Kimi K2.7 Code le…

06:51
2026-08-01
twitter.com
artificial-intelligence

GPT has proved nonsofic groups are exist

OpenAI's GPT 5.6 Sol has reportedly proved the existence of nonsofic groups, a major result in mathematics, according to a post on X. The claim suggests that the AI system has outperformed human mathe…

15:26
2026-07-30
lesswrong.com
large-language-models

Testing LLMs on Undergraduate Music Theory

A test of five modern LLMs on undergraduate music theory found that GPT 5.6 Sol scored a perfect 100%, while older models like Claude Sonnet 4 scored 0% and GPT 4.1 scored 16%, indicating LLMs have su…

09:21
2026-07-29
blog.us.fixstars.com
large-language-models

Can a 2.8T Model Run on a Single Node of Nvidia B300 X8?

Moonshot AI released the 2.8-trillion-parameter Kimi-K3 open-weight model on July 27, 2026, and Fixstars successfully ran inference on a single-node NVIDIA B300 x8 system using SGLang's DCP support, a…

23:28
2026-07-26
ziva.sh
artificial-intelligence

Godot Benchmark 2: Opus 5 > Sol > Terra

Ziva's Godot Benchmark 2 finds Claude Opus 5 is the only model that produces a playable 3D vampire survivor game, scoring 8/10 for world quality and 9/10 for conversation, but costing $65.64 and takin…

20:13
2026-07-22
lesswrong.com
artificial-intelligence

Can an LLM make a feature-length movie on its own?

A filmmaker used LLMs including Claude Fable 5, GPT 5.6 Sol, and Veo 3.1 to create a feature-length adaptation of William Hope Hodgson's book, but deemed the result a failure due to LLMs' poor sense o…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics