Google releases three new Gemini models — but no 3.5 Pro
Google prioritizes efficiency and cybersecurity capabilities with 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber, while delaying 3.5 Pro and signaling a pacing shift in flagship capabilities.
At a glance
- 3.6 Flash price and efficiency: $1.50/1M input tokens and $7.50/1M output tokens; token usage reduced by about 17% versus 3.5 Flash.
- 3.5 Flash-Lite delivers 350 tokens per second and costs $0.30/1M input tokens and $2.50/1M output tokens.
- 3.5 Flash Cyber will be available to governments and trusted partners via a limited pilot.
- 3.5 Pro remains in testing with partners and is not yet broadly released.
The story
Google DeepMind announced Gemini 3.6 Flash, along with 3.5 Flash-Lite and 3.5 Flash Cyber, with a focus on efficiency, lower latency, and reliability for scaling agent-based AI workflows. Google says 3.6 Flash reduces token usage by about 17% and pricing is lower than 3.5 Flash, at $1.50 per 1M input tokens and $7.50 per 1M output tokens.
Gemini 3.5 Flash-Lite is described as the fastest model in the 3.5 family, achieving 350 tokens per second and offering a strong price-to-performance ratio for high-throughput workloads. It is priced at $0.30 per 1M input tokens and $2.50 per 1M output tokens.
Gemini 3.5 Cyber is a specialized cybersecurity-tuned model that Google plans to pilot for limited government and trusted-partner use via CodeMender, highlighting the dual-use nature of such tools. The flagship Gemini Pro 3.5 Pro was not released and remains in testing with partners.