cd /news/artificial-intelligence/ainews-quasi-riemann-hypothesis-open… · home › topics › artificial-intelligence › article
[ARTICLE · art-146596] src=latent.space ↗ pub= topic=artificial-intelligence verified=true sentiment=↑ positive

[AINews] Quasi-Riemann-Hypothesis: OpenAI publishes 722 math papers solving 90 of the top 500 open math problems; “the most significant moment” in >100 years of mathematics

OpenAI published 722 mathematical manuscripts from an unreleased internal frontier model on a public GitHub repo, drawing on an evaluation of roughly 4,000 research problems at an average of about three hours of ChatGPT Pro thinking compute per result. The release includes papers, proof artifacts and selected reasoning summaries, with Sam Altman calling it "a new era of discovery," while the model itself remains unreleased and the claimed results have not been independently verified. Commentators singled out Result 003, the Quasi-Riemann Hypothesis, as somewhere between a Fields Medal result and the biggest result in number theory in 200 years.

read11 min views2 publishedOct 7, 2026
[AINews] Quasi-Riemann-Hypothesis: OpenAI publishes 722 math papers solving 90 of the top 500 open math problems; “the most significant moment” in >100 years of mathematics
Image: Latent Space

Tickets for AIE NYC are selling out soon! See you next week! see past AINews issues for subscriber discounts.

Pour one out for Mistral, who shipped a decent Large 4 “Le Chonk” model on the new 3800 GB300 cluster funded by their recent Series D.

But they were overshadowed by more mathematics results from OpenAI’s internal Navier-Stokes math model - published as a blogpost, repo, and tweet. The best compliment comes from their Navier-Stokes competitor from Anthropic, who despite his personal issues with Anthropic, does not mince words: “It’s obviously the most significant moment in mathematical history.”

This bears some qualification, but most experts seem to agree that it solves many of the top 500 open problems in math.

In particular, Result 003, the Quasi-Riemann Hypothesis, is somewhere between a Fields Medal result and “the biggest result in number theory in 200 years”.

The most astonishing is the how - while Navier-Stokes was done in 88 hours and 10,000 agents, these solutions were 3 hours of ChatGPT Pro on average.

AI News for 10/5/2026-10/6/2026. We checked 12 subreddits, 544 Twitters and no further Discords. AINews’ website lets you search all past issues. As a reminder, AINews is now a section of Latent Space. You can opt in/out of email frequencies!

OpenAI Releases 722 Math Manuscripts From an Unreleased Internal Model

Mistral Large 4 (”Le Chonk”): Launch, Pricing and Contested Evals

  - **Coding benchmarks** : It reports outperforming GLM 5.3 on DeepSWE and Kimi K3 on Terminal-Bench 4 ([Rozière](https://x.com/b_roziere/status/2107470925952344505) ).
  - **Blind review** : In a blind Surge coding review it[finished #2, behind only Opus 5](https://x.com/echen/status/2107504639968940534) .
- **Independent measurements** :
  - **Index gap** : Critics note it trails[GLM-5.3 and even GLM-5.3-Flash on AA’s index](https://x.com/Yuchenj_UW/status/2107468232106078433) .
  - **Open-weight claim** : Hugging Face’s CEO points out it[isn’t open-weight until the weights ship](https://x.com/ClementDelangue/status/2107525319012090301) .

Open-Weight and API Model Releases: Embeddings, Image, Decision Models

  • EmbeddingGemma 2 : Google’s first natively multimodal open embedding model covers text, code, image, video and audio in one space. It is built on Gemma 4 and released under Apache 2.0 (DeepMind ).
    • Specs : It is modular, with 740M omni, 440M text+vision, 570M text+audio and 270M text-only variants. It has Matryoshka dimensions from 768 down to 128, 8,192 context and a reported +14% on MTEB Code (Phil Schmid ).
    • Footprint : It uses roughly 191–567MB of active RAM and handles up to 5.5 minutes of audio or 58 video frames per pass (Google ).
    • Ecosystem : Day-0 support coversllama.cpp ,vLLM ,Ollama andUnsloth . It also runs in the browser on WebGPU at~20–70ms per query .
  • Nano Banana 2.1 : Google’s updated image model is rolling out across the Gemini app, AI Studio, Search and Ads (Google ).
  • Decision models become a product category :
    • OpenAI Decisions API : The public beta runs on GPT-6 Luna and returns predicates, choices or scores. OpenAI says it is up to 10x faster than the Responses API (OpenAI Devs ). Pricing starts at$0.10/M input with no output charges .
    • Perplexity :pplx-decider-v1.1-27b is open weights, costs $0.02/M input and tops the new HF Decision Index v0.3.
    • Independent check on Jev :Vals found Jev matched GPT-6 Astra’s 97.5% on claim verification at about 1/500th the cost. Jev alsoranked last on LegalBench .
    • Skeptic view : Theo argues model-routing use cases are“absolutely useless” for choosing intelligence levels.
  • Other open releases :
    • Ling 3.1 Flash : The model has 560B total and 25B active parameters and scores41 on AA’s index , up from 20. It costs $0.30/$0.90 per million tokens, and weights are coming.
    • Reflection Beam : A Zhihu analysis ofBeam describes a 501B/23B MoE with 23.8T pretraining tokens. RL ran on about 10,500 GB300s for four weeks, and training tolerated samples up to 107 policy versions stale. Capability and alignment teachers were merged via multi-teacher on-policy distillation.
    • Kandinsky 6.0 : The video model ships under an MIT license with synchronized audio andday-0 vLLM-Omni support .
  • Search eval : OpenAI’s built-in web search scores74 on the AA Search Index , 5th among providers, at about $0.05 per task. It is weakest on BrowseComp, where it ranks 13th of 26.

Safety, Control and Eval Integrity

Research, Infrastructure and Developer Tools

  - **Scale AI** : Scale[open-sourced AgentEnv](https://x.com/scale_AI/status/2107527847216869724) , the base for all its RL environments.
  - **Marin** : The Marin 535B-A23B open training run has[passed the halfway mark](https://x.com/percyliang/status/2107502164902031487) .
- **Hardware** :

Industry and Policy

  • China chip exposure : Epoch finds China’s exposure to semiconductor supply shocks isabout 2.7x that of the US . Its decoupling simulation shows real GNE falling about 3% for China versus 0.6% for the US.
  • Chinese AI revenue : A separate Epoch report mapsfive revenue sources for Chinese AI firms . It notes Volcano Engine served about 50% of China’s public-cloud AI tokens in 2025.
  • Qualcomm–Huawei correction : Qualcomm told Yicai that reports linking its deal toHuawei’s LogicFolding technology are untrue . It also disputed reports that it is the net payer.
**Top tweets (by engagement)**

- [Mistral announces Large 4 “Le Chonk”](https://x.com/MistralAI/status/2107457414387622310) (45.6K)
- [OpenAI releases internal-model math results](https://x.com/OpenAI/status/2107596713791767021) (19.0K)
- [Sundar Pichai introduces EmbeddingGemma 2](https://x.com/sundarpichai/status/2107501975671890211) (7.1K)
- [Google AI Studio launches Nano Banana 2.1](https://x.com/GoogleAIStudio/status/2107501303890915550) (7.0K)
- [Anthropic expands Cyber Verification Program](https://x.com/AnthropicAI/status/2107546569654636883) (4.6K)
- [ChatGPT Meetings plugin](https://x.com/ChatGPT/status/2107567930557026653) (3.9K)
- [Integer multiplication faster than n log n in OpenAI’s math repo](https://x.com/AcerFur/status/2107606747972309163) (3.1K)
- [OpenAI Decisions API public beta](https://x.com/OpenAIDevs/status/2107573382229188645) (2.8K)

/r/LocalLlama + /r/localLLM Recap #

1. Local AI Tooling Releases

  • google/embeddinggemma-2 · Hugging Face (Activity: 543):Google DeepMind released google/embeddinggemma-2, a 740M-parameter open multimodal embedding model mapping text/code, images, video, audio, and mixed inputs into a shared 768d space for on-device retrieval/RAG/classification/clustering. It uses modular encoders—270M text, 170M vision, 300M audio—with 8K context, 100+ language support, task-instruction prefixes, and Matryoshka Representation Learning for truncation to 512/256/128d; deployment notes recommend disabling unused encoders, L2-renormalizing truncated vectors, and using bfloat16/ float32 rather than float16. Community links include llama.cpp support PR #30054, ggml-org GGUF weights, and Unsloth GGUF weights. Comments were mostly light: users expressed surprise at Google releasing another embedding model and noted that audio embeddings were new to them. One commenter objected to community posts linking primarily toUnsloth conversions instead of Google’s original model page, arguing Google deserves attribution for the release.
    • llama.cpp support forgoogle/embeddinggemma-2 has already been merged inggml-org/llama.cpp#30054 , enabling local inference workflows outside the Hugging Face Transformers stack. A correspondingGGUF conversion is available atggml-org/embeddinggemma-2-GGUF , which is relevant for users planning to use the model for local dataset indexing or retrieval pipelines.
── more in #artificial-intelligence 4 stories · sorted by recency
── more on @openai 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/ainews-quasi-riemann…] indexed:0 read:11min 2026-10-07 · —