{"slug": "google-expands-gemini-with-3-6-flash-flash-lite-and-gemini-robotics-2", "title": "Google Expands Gemini With 3.6 Flash, Flash-Lite and Gemini Robotics 2", "summary": "Google has expanded its Gemini family with the release of Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and Gemini 3.5 Flash Cyber, alongside the new Gemini Robotics 2 family for embodied control. The Flash models are designed for agentic workflows, with 3.6 Flash offering a 17% reduction in output tokens and 3.5 Flash-Lite delivering up to 350 output tokens per second. Pricing for 3.6 Flash is set at $1.50 per million input tokens and $7.50 per million output tokens, while 3.5 Flash Cyber is available through a limited pilot for governments and trusted partners.", "body_md": "Google is expanding Gemini on two fronts at once: faster, lower-cost models for software and enterprise workflows, and a new robotics family designed for embodied, cross-robot control. The releases include **Gemini 3.6 Flash**, **Gemini 3.5 Flash-Lite**, Gemini 3.5 Flash Cyber, and [ Gemini Robotics 2](https://scalevise.com/resources/gemini-robotics-2-google-deepmind-actual-lineup/) with related embodied-reasoning and on-device variants.\n\nThe clearest immediate enterprise story is the widening choice of models for agentic work. In its [official Gemini Flash announcement](https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber/), Google positions 3.6 Flash as a general workhorse for coding, knowledge work, and multimodal tasks, while 3.5 Flash-Lite is aimed at workloads where response speed and cost efficiency are decisive. The robotics update extends the same broader push beyond software agents into systems that must reason about and act in physical environments.\n\nGemini 3.6 Flash is generally available through Google's developer, enterprise, and consumer channels. Google says it improves on 3.5 Flash for coding, knowledge-work, and multimodal tasks, while producing around **17% fewer output tokens** than 3.5 Flash. That token-efficiency claim matters because output tokens are a material part of both latency and inference spending in multi-step agent workflows.\n\nGoogle lists pricing for Gemini 3.6 Flash at **$1.50 per 1 million input tokens** and **$7.50 per 1 million output tokens**. The company describes the model as offering a lower cost per task, a metric that depends not only on token prices but also on how many tokens a task requires to complete. The supplied release information does not provide a price for 3.5 Flash-Lite, so it should not be inferred from 3.6 Flash pricing.\n\nGemini 3.5 Flash-Lite occupies a different role. Google calls it its fastest and most cost-effective subfamily, with a stated output speed of **350 output tokens per second**. It is intended for high-throughput agentic workflows where an organization may value quick model responses and high request volume over the broader workhorse positioning of 3.6 Flash.\n\n| Release | Primary positioning | Availability or access | Verified detail |\n|---|---|---|---|\n| Gemini 3.6 Flash | Workhorse for coding, knowledge work, and multimodal tasks | General availability across Gemini APIs, Google AI Studio, Android Studio, Antigravity, Gemini Enterprise, and the Gemini app | $1.50 per million input tokens; $7.50 per million output tokens |\n| Gemini 3.5 Flash-Lite | Fast, cost-effective model for high-speed agentic workflows | Available through the Gemini ecosystem, including Google Search integration | Up to 350 output tokens per second |\n| Gemini 3.5 Flash Cyber | Cybersecurity-focused model for vulnerability work | Limited pilot for governments and trusted partners | Integrated with CodeMender for finding and patching vulnerabilities |\n\nThe availability split is significant. 3.6 Flash and 3.5 Flash-Lite are broadly available across the Gemini ecosystem, including the Gemini App, Gemini Enterprise App, Gemini Enterprise Agent Platform, Google AI Studio, Gemini API, and Google Antigravity. By contrast, 3.5 Flash Cyber is being handled through a limited pilot for governments and trusted partners. Google also says Gemini 3.5 Pro is testing with partners, rather than presenting it as broadly available.\n\nFor technical teams, the practical decision is less about selecting a single \"best\" model than matching a model to the economics and risk profile of a workflow:\n\nOrganizations evaluating these options can work with Scalevise on AI architecture, [workflow automation](https://scalevise.com/resources/ai-workflow-automation/), and integration plans that connect model selection to real operational requirements.\n\nGoogle DeepMind's robotics release is not simply another text or code model. **Gemini Robotics 2** is part of an embodied AI family that also includes Gemini Robotics ER 2 and Gemini Robotics On-Device 2. Google describes the family as enabling whole-body intelligence and cross-embodiment control for humanoids and other robots.\n\nThat framing points to a different deployment challenge from conventional enterprise AI. A software agent can operate through APIs and tools, while a robot has to connect perception, reasoning, movement, and the constraints of a physical body. Cross-embodiment control is especially relevant because it suggests a goal of applying intelligence across more than one robot form factor, rather than limiting a system to a single machine design.\n\nThe robotics offering is live, with a dedicated DeepMind page showing demonstrations and an early-access waitlist. The supplied material does not establish broad general availability or commercial pricing for Gemini Robotics 2, Gemini Robotics ER 2, or Gemini Robotics On-Device 2. For businesses, that makes the update an important indicator of Google's direction in physical AI, but not yet a basis for assuming that robotics capabilities can be deployed under the same access model as the generally available Flash releases.\n\nTogether, the releases show Google building a more segmented Gemini portfolio. The Flash models address common enterprise concerns around speed, token use, tool-enabled workflows, and deployment reach. The robotics family addresses a longer-horizon need: bringing multimodal reasoning into systems that interact directly with the physical world. The common thread is an effort to make Gemini useful across increasingly varied forms of agency, from software workflows to robots.\n\n**What is Gemini 3.6 Flash?**\n\nGemini 3.6 Flash is Google's generally available workhorse model for coding, knowledge-work, and multimodal tasks. Google prices it at $1.50 per 1 million input tokens and $7.50 per 1 million output tokens.\n\n**How is Gemini 3.5 Flash-Lite different from Gemini 3.6 Flash?**\n\nGoogle positions 3.5 Flash-Lite as its fastest and most cost-effective subfamily for high-speed agentic workflows, with up to 350 output tokens per second. Gemini 3.6 Flash is positioned more broadly for coding, knowledge work, and multimodal tasks.\n\n**Where are Gemini 3.6 Flash and 3.5 Flash-Lite available?**\n\nGoogle lists general availability through the [Gemini App](https://scalevise.com/resources/gemini/), Gemini Enterprise App, Gemini Enterprise Agent Platform, Google AI Studio, Gemini API, and Google Antigravity. Google also identifies Google Search integration for 3.5 Flash-Lite.\n\n**Is Gemini Robotics 2 broadly available?**\n\nThe Gemini Robotics 2 family is live and has an early-access waitlist, but the supplied information does not confirm broad general availability or commercial pricing.\n\nGoogle's Gemini updates create a clearer division of labor across AI workloads. Gemini 3.6 Flash and 3.5 Flash-Lite provide broadly available options for enterprise and developer workflows with different performance priorities, while Gemini Robotics 2 extends the program toward embodied intelligence. The key near-term question is how organizations translate that expanding model range into dependable, appropriately governed software and robotics deployments.", "url": "https://wpnews.pro/news/google-expands-gemini-with-3-6-flash-flash-lite-and-gemini-robotics-2", "canonical_source": "https://dev.to/alifar/google-expands-gemini-with-36-flash-flash-lite-and-gemini-robotics-2-3npj", "published_at": "2026-08-01 12:30:30+00:00", "updated_at": "2026-08-01 12:40:32.607419+00:00", "lang": "en", "topics": ["artificial-intelligence", "large-language-models", "ai-products", "ai-infrastructure", "ai-agents"], "entities": ["Google", "Gemini 3.6 Flash", "Gemini 3.5 Flash-Lite", "Gemini 3.5 Flash Cyber", "Gemini Robotics 2", "Google DeepMind", "Scalevise", "CodeMender"], "alternates": {"html": "https://wpnews.pro/news/google-expands-gemini-with-3-6-flash-flash-lite-and-gemini-robotics-2", "markdown": "https://wpnews.pro/news/google-expands-gemini-with-3-6-flash-flash-lite-and-gemini-robotics-2.md", "text": "https://wpnews.pro/news/google-expands-gemini-with-3-6-flash-flash-lite-and-gemini-robotics-2.txt", "jsonld": "https://wpnews.pro/news/google-expands-gemini-with-3-6-flash-flash-lite-and-gemini-robotics-2.jsonld"}}