Gemini 3.6 Flash lands — and the Pro everyone wants still isn't here.
On July 21, 2026, Google shipped three new Gemini models: 3.6 Flash (the new everyday workhorse, already in the Gemini app), 3.5 Flash-Lite (the speed-and-price model, rolling into Google Search), and 3.5 Flash Cyber (a security specialist you can't have). What Google didn't ship is the one everyone's waiting for: Gemini 3.5 Pro. Here's what actually changed for you, what the confusing names mean, and where the flagship is.
01 Three models, one minute
| Model | What it's for | Who gets it |
|---|---|---|
| Gemini 3.6 Flash | The new default workhorse — better coding, knowledge work, and multimodal answers, with less rambling | Everyone, in the Gemini app — plus developers (AI Studio, Android Studio, Antigravity) and Gemini Enterprise |
| Gemini 3.5 Flash-Lite | Speed and volume: 350 tokens/second, Google's cheapest 3.5-class model | Developers today; rolling out inside Google Search |
| Gemini 3.5 Flash Cyber | Fine-tuned to find and fix software security vulnerabilities, inside Google's CodeMender agent | Governments and trusted partners only — limited-access pilot |
02 What 3.6 Flash actually improves
The headline isn't raw intelligence — it's efficiency. Google says 3.6 Flash uses 17% fewer output tokens than 3.5 Flash on the Artificial Analysis index while scoring better on coding (DeepSWE: 49% vs 37%), computer-use tasks (OSWorld-Verified: 83.0% vs 78.4%), and knowledge work. In practice that means faster, less padded answers that take fewer steps to finish multi-part tasks.
For developers it's directly cheaper — $1.50 per million input tokens and $7.50 per million output tokens, less than 3.5 Flash. For everyone else, token efficiency shows up as speed: shorter chains of reasoning between your question and a finished answer. Note those prices are developer API rates, not what Gemini app subscribers pay.
One more quiet upgrade: computer use is now a built-in tool — the model can drive on-screen interfaces, which is the plumbing behind agents that click and type for you. We covered that capability when it debuted in 3.5 Flash; 3.6 does it measurably better.
03 The elephant: where's 3.5 Pro?
Google's flagship Pro model was last updated in February. Since then OpenAI shipped GPT-5.5 and GPT-5.6, and Anthropic shipped Opus 4.8, Sonnet 5, and Fable 5. In Tuesday's post Google says 3.5 Pro is "currently testing with partners" and will be broadly available "as soon as it's ready" — and that pre-training for Gemini 4, its "most ambitious" run yet, has already started.
The honest read: Flash models are Google's volume play — fast, cheap, and now genuinely good. But if you're choosing a tool for the hardest reasoning work today, Google's best answer is five months old. If your work leans on frontier reasoning, our which-AI-for-which-job lesson stays the guide until Pro lands.
What to do with this
- In the Gemini app, just keep working — 3.6 Flash is arriving as the everyday model, no toggle hunt needed
- Notice answer style: if Gemini suddenly feels snappier and less wordy this week, that's 3.6 Flash
- Building on the API? 3.6 Flash is a drop-in that's cheaper per task than 3.5 Flash — test your agent flows on it
- Don't wait on Flash Cyber — it's a government/partner pilot, not a consumer feature
- Watching for 3.5 Pro? So are we — we'll ship the lesson the day it's real