Google announced a flagship it couldn't ship — and called the cheap one the win
Gemini 3.5 Flash is out. Gemini 3.5 Pro is a 'next month'. Notice which one got the keynote.
The answer
Google shipped Gemini 3.5 Flash at I/O 2026; Gemini 3.5 Pro was never released.
At I/O on 19 May 2026, Google launched Gemini 3.5 Flash and made it the default for the Gemini app and AI Mode in Search globally, with the model also in the Gemini API. Flash is genuinely useful. Being the default in Google Search's AI layer — reportedly past a billion monthly users — is a distribution advantage no rival can match on day one. All of that is real, and none of it is the interesting part of the announcement.
The flagship that wasn't there
Gemini 3.5 Pro got the flagship billing — the frontier model that's supposed to signal Google's challenge to OpenAI's o-series. And then the announcement said the quiet part: 3.5 Pro is 'already being used internally, and we look forward to rolling it out next month.' That is the corporate-keynote equivalent of 'the cheque's in the post.' Here's the harder tell: Google put a name on the flagship and nothing else — no benchmark scores, no context-window figure, no price. A real product launch ships numbers; a placeholder ships a date. 'Next month' is the date; everything that would let you evaluate the model is missing. The tech press largely reprinted the name and moved on.
Announcing a frontier model you cannot actually hand to users is a standard competitive move: it freezes developer conversations ('let's wait and see what Google's Pro can do') and steals a few news cycles from whoever shipped this week. It only works if 'next month' arrives. Every week it doesn't, the tactic calcifies into a credibility problem. Developers have long memories for vaporware, and in a market where Claude, GPT, and Llama derivatives are shipping aggressively, 'it's coming' isn't a strategy — it's a countdown.
Google introduced Gemini 3.5 Flash at I/O 2026, positioning it as beating the previous Pro tier on key benchmarks — while simultaneously announcing Gemini 3.5 Pro as a model that would arrive 'next month'.
The price the keynote skipped past
Here is the bit that did not make the keynote slides. Official API pricing for Gemini 3.5 Flash is $1.50/M input and $9/M output (cached input $0.15/M). Google's only pricing line is relative — Flash costs 'less than half' of other frontier models — which is a comparison to flagships, not a discount on the old Flash. Notice what's missing: Google never put the prior Gemini 3 Flash price next to the new one. For the budget tier, $9 per million output tokens is not nothing, and output is where agentic apps burn tokens. 'Cheaper than a flagship' and 'cheap' are not the same claim, and the keynote leaned on the first to imply the second. If you are building anything at scale on Gemini Flash, re-price your unit economics against the published rate card before it surprises you in a billing cycle.
The benchmark you should verify
Google cited 76.2% on Terminal-Bench 2.1 as the headline capability number for Flash. Terminal-Bench is a legitimate, non-trivial evaluation of long-horizon coding and shell interaction — not a toy benchmark — which makes the claim worth taking seriously. It also makes it worth verifying before you treat it as fact. Artificial Analysis and LMSYS are running independent evaluations; their results will land within weeks. Until then, file 76.2% as 'Google says so', because that is what it currently is. The claim may well hold. Waiting for one independent confirmation is not scepticism — it is basic epistemic hygiene in an industry that has a long track record of benchmark-optimised figures.
Set the Gemini 3.5 Flash story against the full context:
| What Google said | What to independently check | |
|---|---|---|
| Capability | Beats Gemini 3.1 Pro on coding/agentic/multimodal | Artificial Analysis / LMSYS independent evals |
| Benchmark | 76.2% Terminal-Bench 2.1 | Has any third party reproduced this? |
| Price | 'Less than half' the cost of frontier models | $1.50/M in, $9/M out — verify on the rate card |
| 3.5 Pro | 'Used internally,' shipping 'next month' | Has it actually shipped, with real numbers? |
Each row has a verification step. Do those four checks and you know more than the keynote told you.
3.5 Flash is now the default model for the Gemini app and AI Mode in Search globally. We're also hard at work on 3.5 Pro. It's already being used internally, and we look forward to rolling it out next month.
What to watch
Watch three things: whether 3.5 Pro actually ships 'next month' with real numbers attached (the credibility clock is ticking); whether independent benchmarks confirm the Terminal-Bench 2.1 figure; and whether developer adoption of Flash holds at $9/M output. If Pro slips and benchmarks soften, this announcement ages as a strong distribution play dressed as a capability leap. If Pro ships on time and the numbers hold, it ages as one of the sharper I/O moments in recent years. The receipts will be in within a quarter.
Frequently asked questions
Can I use Gemini 3.5 Pro now?
Why does Flash's pricing matter for developers?
Is the 76.2% Terminal-Bench 2.1 claim credible?
Did Google share any specs for Gemini 3.5 Pro?
Is Google's distribution advantage real?
Sources
- Gemini 3.5: frontier intelligence with action — Google, 19 May 2026
- Google introduces Gemini 3.5 Flash at I/O 2026 — a faster, cheaper model for AI agents and coding — MarkTechPost, 20 May 2026
- Google Search's I/O 2026 updates: AI agents and more — Google, 19 May 2026
- 100 things we announced at Google I/O 2026 — Google, 19 May 2026