Google Ships Gemini 3.6 Flash, Flash-Lite, Cyber

Get the Tech newsletter
Daily tech — startups, AI labs, chips, the launches that shape the next decade. Free.
- Google released three new Gemini models — 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber — aimed at production AI agent workflows.
- Gemini 3.6 Flash consumes 17% fewer output tokens than 3.5 Flash on the Artificial Analysis Index, with gains on DeepSWE (49% vs 37%), MLE Bench (63.9% vs 49.7%), and OSWorld-Verified (83.0% vs 78.4%), priced at $1.50/$7.50 per million input/output tokens.
- Gemini 3.5 Flash-Lite runs at 350 output tokens per second at $0.30/$2.50 per million tokens, outperforming the older 3.1 Flash-Lite on Terminal-Bench 2.1 (54% vs 31%) and even beating 3 Flash on SWE-Bench Pro (54.2% vs 49.6%).
- Gemini 3.5 Flash Cyber is fine-tuned for finding and patching security vulnerabilities and will be restricted to governments and trusted partners via the CodeMender agent as a limited-access pilot.
- Google said 3.6 Flash ships with enhanced CBRN and cyber-offense safety safeguards making it "substantially more resistant to jailbreaks," while Gemini 3.5 Pro is testing with partners and Gemini 4 pre-training has begun.
Why it matters: Developers building production AI agents get a meaningfully cheaper tier — 3.6 Flash at $1.50/$7.50 per million tokens with 17% fewer tokens consumed per task — while Google's restricted-access 3.5 Flash Cyber signals frontier defensive AI tooling will sit behind vetted-partner walls rather than the open market.

