Google releases Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber, but no 3.5 Pro
Google is pushing cost efficiency and speed for AI agents with the new Flash lineup, while reserving a limited pilot for its cybersecurity-focused 3.5 Cyber and delaying a flagship 3.5 Pro. This shapes enterprise planning for cost and safety in AI deployments.
At a glance
- Gemini 3.6 Flash costs: $1.50 per 1M input tokens and $7.50 per 1M output tokens, with 17% fewer tokens than 3.5 Flash.
- 3.5 Flash-Lite delivers 350 tokens per second and prices at $0.30 per 1M input tokens and $2.50 per 1M output tokens.
- 3.5 Cyber is being rolled out as a limited pilot for governments and trusted partners.
- 3.5 Pro is in testing with partners but has not been released broadly.
The story
Google announced three new Gemini models today: Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and Gemini 3.5 Flash Cyber. The company described 3.6 Flash as an improvement over the 3.5 line with better coding and multimodal capabilities, and it runs with lower token usage and a reduced price point.
Gemini 3.5 Flash-Lite is pitched as the fastest model in the 3.5 family, capable of 350 output tokens per second, and priced at $0.30 per 1M input tokens and $2.50 per 1M output tokens. In OSWorld tests, 3.5 Flash-Lite showed strong performance gains on throughput and real-world tasks. 3.5 Flash-Lite is intended for high-throughput production workloads.
Gemini 3.5 Cyber will be released as a limited pilot focused on cybersecurity use cases and will be available only to governments and trusted partners via CodeMender. Google also noted that 3.5 Pro—its planned top-tier model—remains in testing with unnamed partners and will be released when ready.