{"slug": "google-tests-new-gemini-4-pro-checkpoints-early-outputs", "title": "Google tests new Gemini 4 Pro checkpoints, early outputs", "summary": "Google has begun internal testing of early Gemini 4 Pro checkpoints under the code name \"argon,\" according to a leak from Lentils, with developer Bee posting a demo on X of a monographic website the model generated in fourteen minutes. Bee said the output was \"so good\" and that Google had done a great job with the design, addressing Gemini's long-standing weakness in frontend design taste. Google is also testing Projects for the Gemini desktop app, which will work like the workspaces in ChatGPT or Claude with dedicated hubs for reference files and custom instructions.", "body_md": "Google has been steadily dropping newer Flash models over the past few months, but hardcore users have only been waiting for one thing: a new frontier-class model. The recent [Gemini 3.8 Flash release](https://www.testingcatalog.com/google-releases-gemini-3-8-flash-and-flash-cyber/) brought some interesting coding improvements, but it would be a disservice to call it a breakthrough model.\n\nGemini models have long had a reputation for producing awkward, unoriginal web designs when asked to create frontend code. That might finally be on its way out, if early test results of what appears to be [Gemini 4 Pro](https://www.testingcatalog.com/google-prepares-gemini-app-for-avatars-plugins-and-gemini-4/) are anything to go by.\n\nThe leak from Lentils was among the first to mention it a few days ago, saying Google had begun testing early checkpoints of Gemini 4 Pro within the company under the code name *\"argon.\"*\n\nSince then, others have also shared their early Gemini 4 Pro test results on X. A demonstration posted by developer Bee on X featuring a fully monographic website that looked extremely clean was pretty impressive.\n\nAlthough the model took fourteen minutes to produce the page, the final result had nothing whatever in common with the robotic templates which we typically receive from AI tools. Bee said that he was surprised at how clean and well-polished the monographic website had become *\"it looks so good\"*, and mentioned that Google had done a great job with the design, the earlier problems regarding design taste now appearing to have been resolved.\n\nThe audio effects used to create the sound of sketching on canvas paper are also a neat touch that adds to the interface's tactile feel.\n\nAnother AI benchmarking account also shared Gemini 4 Pro's output in an attempt to create an Xbox controller SVG. We've seen similar tests from Gemini models in the past, but none have had the same level of polish as this one. So it's almost certain this is Gemini 4 Pro's work of art.\n\nWay back in July, Logan Kilpatrick [hinted](https://x.com/OfficialLoganK/status/2067646759078375651?ref=testingcatalog.com) that the team is cooking with *\"Gemini 3.5 Pro,\"* but it looks like Google may not have been too excited about the results at the time. With all the progress made since then, it only makes sense that the team would skip the 3.5 moniker and go straight to 4 Pro. Plus, with these early results, it might be worth the big jump instead of a point release.\n\nThat said, while Gemini 4.0 Pro is still in the oven, Google is also working on enhancements for the Gemini desktop app, which [we spotted in testing](https://www.testingcatalog.com/google-tests-proper-projects-for-the-gemini-desktop-app/). In short, Projects will work a lot like the workspaces in ChatGPT or Claude. You’ll have dedicated hubs for reference files and custom instructions, so each new conversation uses that shared context instead of making you start over every time.\n\nThese tests also come at an interesting time, as we're seeing some of the first signs of what could be Anthropic's [upcoming Fable 5.2 model](https://www.testingcatalog.com/anthropic-tests-fable-5-2-and-opus-5-5-ahead-of-the-release/). So it'll be interesting to see what Google has been working on for the past few months and how it compares to Anthropic's next frontier model. \n\nThe results from these current tests suggest that Gemini 4 Pro could possibly have a real chance of becoming a worthy frontier-class model from Google. But we'll have to wait and see how it fares in head-to-head comparisons with Astra, Grok 4.7, and, of course, Fable 5.2 to see how it stacks up against the competition.", "url": "https://wpnews.pro/news/google-tests-new-gemini-4-pro-checkpoints-early-outputs", "canonical_source": "https://www.testingcatalog.com/gemini-4-pro-frontend-ui-taste-leak/", "published_at": "2026-09-23 07:53:00+00:00", "updated_at": "2026-09-23 07:54:52.656443+00:00", "lang": "en", "topics": ["large-language-models", "generative-ai", "ai-products", "ai-tools"], "entities": ["Google", "Gemini 4 Pro", "Gemini 3.8 Flash", "Lentils", "Bee", "Logan Kilpatrick", "Anthropic", "Fable 5.2"], "alternates": {"html": "https://wpnews.pro/news/google-tests-new-gemini-4-pro-checkpoints-early-outputs", "markdown": "https://wpnews.pro/news/google-tests-new-gemini-4-pro-checkpoints-early-outputs.md", "text": "https://wpnews.pro/news/google-tests-new-gemini-4-pro-checkpoints-early-outputs.txt", "jsonld": "https://wpnews.pro/news/google-tests-new-gemini-4-pro-checkpoints-early-outputs.jsonld"}}