Your Best Prompt Is a Decade of Expertise
On July 21, mathematician Terence Tao published a raw ChatGPT transcript showing how he used the model to analyze a counterexample to the Jacobian Conjecture, an algebra problem open since 1939, and G…
On July 21, mathematician Terence Tao published a raw ChatGPT transcript showing how he used the model to analyze a counterexample to the Jacobian Conjecture, an algebra problem open since 1939, and G…
Alibaba released Qwen3.8-Max, a 2.4-trillion-parameter sparse Mixture-of-Experts model, on August 3, claiming it outperforms GPT-5.6 Sol on key coding benchmarks and promising full open weights next w…
Thinking Machines Lab released Inkling on July 15, a 975-billion-parameter Mixture-of-Experts model with 41B active parameters, open-weight under Apache 2.0, from Mira Murati's $12B startup. The model…
Alibaba has released Qwen3.8-Max, a 2.4-trillion-parameter mixture-of-experts model with 95 billion active parameters per token and a one-million-token context window. The company claims the model ope…
Alibaba's Qwen3.8-Max, a 2.4-trillion-parameter open-weights model, launched and ranks behind only Claude Fable 5 and select Opus variants on Arena.AI, according to The Verge. The EU AI Act's transpar…
The White House will host a meeting Tuesday with AI executives to discuss a voluntary system of government review of the most powerful AI models, an administration official said. The framework, outlin…
A developer accepted a friend's dare to build a Pass the Pigs app during an apéritif, using an adversarial AI development loop. The plan was written by Codex and reviewed by Claude Fable 5, with Herme…
Alibaba Group Holding Ltd. shares rose as much as 6% in Hong Kong trading on Monday after the company unveiled Qwen3.8-Max, a 2.4 trillion parameter mixture-of-experts model that Alibaba claims trails…
In an informal experiment, five AI models—GPT-5.6 Sol, Claude Fable 5, Gemini 3.6 Flash, Kimi Instant, and Grok 4.5 Fast—each independently selected Bitcoin as the single cryptocurrency to hold for fi…
Chinese AI startup DeepSeek launched V4-Flash on Friday, the cheapest major AI model to operate at an average cost of 3 cents per test, according to research firm Artificial Analysis, which found it c…
Alibaba's Qwen team released Qwen3.8-Max, a 2.4-trillion-parameter mixture-of-experts model, now generally available via API with open weights promised next week. The model scores 86.6 on Terminal-Ben…
Steve Yegge, a software engineer and creator of the game Wyvern, announced that his new closed-source coding harness, Wheelhouse, is bespoke and not reusable, and he predicts that all harnesses will s…
Alibaba released Qwen3.8-Max, a 2.4-trillion-parameter model with 95 billion active parameters, claiming it outperforms OpenAI's GPT-5.6 Sol and Anthropic's Claude Fable 5 on several benchmarks includ…
Alibaba launched Qwen3.8-Max, a 2.4 trillion parameter AI model, which immediately became the highest-ranking Chinese text model on the Arena.AI leaderboard, trailing only Anthropic's Claude Fable 5 a…
Mathematician Levent Alpöge announced on X that he used Anthropic's Claude Fable 5 AI model to find a counterexample to the Jacobian conjecture, a problem posed by Ott-Heinrich Keller in 1939. The dis…
Alibaba's Qwen team unveiled Qwen3.8-Max-Preview, a 2.4 trillion-parameter multimodal model, at the World Artificial Intelligence Conference in Shanghai on July 19, claiming it trails only Anthropic's…
Alibaba released Qwen3.8-Max-Preview, a flagship AI model with 2.4 trillion parameters, claiming it trails only Anthropic's Claude Fable 5 in performance, and its shares rose 5.4% in Hong Kong. The mo…
Alibaba has introduced Qwen3.8-Max, its largest AI model with 2.4 trillion parameters, claiming it matches the performance of top models like Anthropic's Claude Fable 5. The release underscores Alibab…
Alibaba's Qwen team released Qwen3.8-Max, a 2.4 trillion parameter multimodal AI model that trails only Anthropic's Claude Fable 5 in internal benchmarks, with open weights scheduled for release next …
A Second Look Fellowship replication of single-forward-pass evals found that Fable 5, Opus 5, and GPT-5.6-Sol can perform 2-hop and 3-hop latent reasoning without chain-of-thought, with GPT-5.6-Sol im…