Google Expands Gemini With 3.6 Flash, Flash-Lite and Gemini Robotics 2
Google's latest Gemini releases span lower-cost, high-speed models for AI workflows and a robotics family focused on whole-body intelligence.
Google is expanding Gemini on two fronts at once: faster, lower-cost models for software and enterprise workflows, and a new robotics family designed for embodied, cross-robot control. The releases include Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, Gemini 3.5 Flash Cyber, and Gemini Robotics 2 with related embodied-reasoning and on-device variants.
The clearest immediate enterprise story is the widening choice of models for agentic work. In its official Gemini Flash announcement, Google positions 3.6 Flash as a general workhorse for coding, knowledge work, and multimodal tasks, while 3.5 Flash-Lite is aimed at workloads where response speed and cost efficiency are decisive. The robotics update extends the same broader push beyond software agents into systems that must reason about and act in physical environments.
A broader Gemini stack for agentic workloads
Gemini 3.6 Flash is generally available through Google's developer, enterprise, and consumer channels. Google says it improves on 3.5 Flash for coding, knowledge-work, and multimodal tasks, while producing around 17% fewer output tokens than 3.5 Flash. That token-efficiency claim matters because output tokens are a material part of both latency and inference spending in multi-step agent workflows.
Google lists pricing for Gemini 3.6 Flash at $1.50 per 1 million input tokens and $7.50 per 1 million output tokens. The company describes the model as offering a lower cost per task, a metric that depends not only on token prices but also on how many tokens a task requires to complete. The supplied release information does not provide a price for 3.5 Flash-Lite, so it should not be inferred from 3.6 Flash pricing.
Gemini 3.5 Flash-Lite occupies a different role. Google calls it its fastest and most cost-effective subfamily, with a stated output speed of 350 output tokens per second. It is intended for high-throughput agentic workflows where an organization may value quick model responses and high request volume over the broader workhorse positioning of 3.6 Flash.
| Release | Primary positioning | Availability or access | Verified detail |
|---|---|---|---|
| Gemini 3.6 Flash | Workhorse for coding, knowledge work, and multimodal tasks | General availability across Gemini APIs, Google AI Studio, Android Studio, Antigravity, Gemini Enterprise, and the Gemini app | $1.50 per million input tokens; $7.50 per million output tokens |
| Gemini 3.5 Flash-Lite | Fast, cost-effective model for high-speed agentic workflows | Available through the Gemini ecosystem, including Google Search integration | Up to 350 output tokens per second |
| Gemini 3.5 Flash Cyber | Cybersecurity-focused model for vulnerability work | Limited pilot for governments and trusted partners | Integrated with CodeMender for finding and patching vulnerabilities |
The availability split is significant. 3.6 Flash and 3.5 Flash-Lite are broadly available across the Gemini ecosystem, including the Gemini App, Gemini Enterprise App, Gemini Enterprise Agent Platform, Google AI Studio, Gemini API, and Google Antigravity. By contrast, 3.5 Flash Cyber is being handled through a limited pilot for governments and trusted partners. Google also says Gemini 3.5 Pro is testing with partners, rather than presenting it as broadly available.
For technical teams, the practical decision is less about selecting a single "best" model than matching a model to the economics and risk profile of a workflow:
- Gemini 3.6 Flash is the general option for tasks that combine coding, reasoning-oriented knowledge work, and multimodal inputs.
- Gemini 3.5 Flash-Lite targets high-speed, high-volume agent operations.
- Gemini 3.5 Flash Cyber is a specialized capability with constrained access, reflecting the higher-risk nature of vulnerability discovery and remediation.
Organizations evaluating these options can work with Scalevise on AI architecture, workflow automation, and integration plans that connect model selection to real operational requirements.
Gemini Robotics 2 takes Gemini into embodied AI
Google DeepMind's robotics release is not simply another text or code model. Gemini Robotics 2 is part of an embodied AI family that also includes Gemini Robotics ER 2 and Gemini Robotics On-Device 2. Google describes the family as enabling whole-body intelligence and cross-embodiment control for humanoids and other robots.
That framing points to a different deployment challenge from conventional enterprise AI. A software agent can operate through APIs and tools, while a robot has to connect perception, reasoning, movement, and the constraints of a physical body. Cross-embodiment control is especially relevant because it suggests a goal of applying intelligence across more than one robot form factor, rather than limiting a system to a single machine design.
The robotics offering is live, with a dedicated DeepMind page showing demonstrations and an early-access waitlist. The supplied material does not establish broad general availability or commercial pricing for Gemini Robotics 2, Gemini Robotics ER 2, or Gemini Robotics On-Device 2. For businesses, that makes the update an important indicator of Google's direction in physical AI, but not yet a basis for assuming that robotics capabilities can be deployed under the same access model as the generally available Flash releases.
Together, the releases show Google building a more segmented Gemini portfolio. The Flash models address common enterprise concerns around speed, token use, tool-enabled workflows, and deployment reach. The robotics family addresses a longer-horizon need: bringing multimodal reasoning into systems that interact directly with the physical world. The common thread is an effort to make Gemini useful across increasingly varied forms of agency, from software workflows to robots.
Frequently Asked Questions
What is Gemini 3.6 Flash?
Gemini 3.6 Flash is Google's generally available workhorse model for coding, knowledge-work, and multimodal tasks. Google prices it at $1.50 per 1 million input tokens and $7.50 per 1 million output tokens.
How is Gemini 3.5 Flash-Lite different from Gemini 3.6 Flash?
Google positions 3.5 Flash-Lite as its fastest and most cost-effective subfamily for high-speed agentic workflows, with up to 350 output tokens per second. Gemini 3.6 Flash is positioned more broadly for coding, knowledge work, and multimodal tasks.
Where are Gemini 3.6 Flash and 3.5 Flash-Lite available?
Google lists general availability through the Gemini App, Gemini Enterprise App, Gemini Enterprise Agent Platform, Google AI Studio, Gemini API, and Google Antigravity. Google also identifies Google Search integration for 3.5 Flash-Lite.
Is Gemini Robotics 2 broadly available?
The Gemini Robotics 2 family is live and has an early-access waitlist, but the supplied information does not confirm broad general availability or commercial pricing.
Conclusion
Google's Gemini updates create a clearer division of labor across AI workloads. Gemini 3.6 Flash and 3.5 Flash-Lite provide broadly available options for enterprise and developer workflows with different performance priorities, while Gemini Robotics 2 extends the program toward embodied intelligence. The key near-term question is how organizations translate that expanding model range into dependable, appropriately governed software and robotics deployments.