Reviews · JULY 22, 2026
Google Ships Gemini 3.6 Flash and 3.5 Flash-Lite, Gates a Cyber Variant — Pro Still Missing
Google DeepMind's July 21 Flash refresh cuts output pricing to $7.50 per million tokens and lifts DeepSWE from 37% to 49%, while the promised 3.5 Pro slips further and Gemini 4 pre-training begins.
Google DeepMind released three Gemini models on July 21, 2026, and pointedly didn't release the one everyone was waiting for. Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and a gated Gemini 3.5 Flash Cyber variant all shipped together; Gemini 3.5 Pro, promised for this window, didn't. Tulsee Doshi, Google's Senior Director of Product Management, framed the launch around efficiency and confirmed that Gemini 4 is now in pre-training. The framing is doing a lot of work.
On the numbers, 3.6 Flash is a genuine cadence upgrade. Output pricing drops from $9.00 to $7.50 per million tokens, input holds at $1.50, and Artificial Analysis clocks it consuming 17% fewer output tokens than 3.5 Flash on identical workloads, with reductions of up to 65% on Datacurve's DeepSWE benchmark. DeepSWE pass@1 climbs from 37% to 49%. OSWorld-Verified computer-use moves from 78.4% to 83.0%. GDPval-AA lands at 1,421, up from 1,349. The knowledge cutoff jumps eighteen months, from January 2025 to March 2026. Context window stays at 1,000,000 input tokens with a 64,000-token output ceiling.
Flash-Lite is the more interesting shape. At $0.30 per million input and $2.50 per million output, running at 350 output tokens per second on Artificial Analysis, it posts 54% on Terminal-Bench 2.1 (up from 3.1 Flash-Lite's 31%) and 72.2% on GDM-MRCR v2. That's a serious agent-tier cheap model.
Then there's Flash Cyber, gated to vetted security partners and pitched by TNW as a direct answer to Anthropic's Mythos and a companion to Google's own CodeMender agent. Shipping a cybersecurity SKU behind a waitlist while the flagship reasoning model slips is a revealing sequencing choice.
Because the missing piece is Pro. Rebecca Bellan at TechCrunch, working from Bloomberg's earlier reporting, notes that OpenAI has moved from GPT-5.5 to GPT-5.6 and Anthropic now fields both Opus 4.8 and Sonnet 5 (with Claude Opus 4.6 still in wide deployment) during the same stretch in which Google's frontier tier has been conspicuously quiet. A Flash refresh, a Lite refresh, and a vertical variant are what a lab ships when the headline model isn't ready and the quarter still needs a release.
The vibes around Gemini 4 pre-training are meant to absorb that. They probably will, in the short term. Google has spent 2026 running a cadence-heavy release calendar into a market where competitors are shipping capability jumps, and the gap between what a Flash-tier release can carry and what a Pro-tier absence signals is now wide enough that the absence is itself the announcement. The efficiency story is real. It's also what you lead with when you don't have the other story yet.
Sources
- https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber/
- https://techcrunch.com/2026/07/21/google-releases-three-new-gemini-models-but-no-3-5-pro/
- https://gcn.com/google-launches-gemini-flash-cybersecurity-model/19924/
- https://9to5google.com/2026/07/21/gemini-3-6-flash-launch/
- https://thenextweb.com/news/google-gemini-3-6-flash-flash-cyber-launch