Blogerroom logoBlogerroom
AI
AI

Google Ships 3 New Gemini Models, Still No 3.5 Pro

AB
Mr. Aayush BhattJuly 23, 20265 min read
๐ŸŒ Language

Google Ships 3 New Gemini Models, Still No 3.5 Pro

Google shipped three new Gemini Flash models on July 21, but its flagship 3.5 Pro remains stuck in testing, weeks past deadline.

Three new Gemini models shipped on July 21, 2026. The one everyone had actually been waiting for wasn't among them. Google DeepMind released Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and Gemini 3.5 Flash Cyber in a single announcement, real, shipping products with real benchmark gains. But Gemini 3.5 Pro, the flagship model Google promised at its I/O conference back in May with a June delivery target, is still nowhere to be found, weeks past its own deadline.

Google's official line, delivered by Gemini product lead Tulsee Doshi, is that 3.5 Pro is currently testing with partners and will ship broadly as soon as it's ready. That's the same non-answer Google gave when the model first missed its June target. The company has now gone from "coming next month" to "no specific date at all," which is a meaningfully different kind of delay.

Three Ships Left the Dock, One Stayed Behind

Gemini 3.6 Flash is the headline release of the three, positioned as Google's workhorse model and a direct replacement for 3.5 Flash. It scores 49% on the DeepSWE coding benchmark, up sharply from 37% for its predecessor, and 83.0% on OSWorld-Verified, a computer-use benchmark, up from 78.4%. According to Google's own figures, referencing the Artificial Analysis Index, the model uses 17% fewer output tokens than 3.5 Flash while getting cheaper: $1.50 per million input tokens and $7.50 per million output tokens, down from $9 per million output tokens previously.

Gemini 3.5 Flash-Lite is the speed play, running at 350 output tokens per second and outperforming the larger 3 Flash on specific benchmarks, including SWE-Bench Pro, where it scores 54.2% against 3 Flash's 49.6%. The third release, 3.5 Flash Cyber, is narrower still: a model specifically tuned to find and fix cybersecurity vulnerabilities, restricted for now to governments and trusted partners as part of a limited pilot program, with what Google describes as enhanced safeguards against cyber and CBRN misuse.

All three are live today across the Gemini app, Google AI Studio, the API, Google Antigravity, and Android Studio. That's a real, substantive product release by any normal standard.

The Model Everyone Actually Wanted Is Still Missing

None of that changes the shape of the story. Google announced Gemini 3.5 Pro at I/O 2026 in May with a stated June launch target. Blogerroom covered that first missed deadline in detail on July 1, when the model was still described as being in limited preview. Seven weeks later, the situation has barely moved: Pro is still testing with partners, with no committed release window, even as three other models shipped cleanly around it.

According to Bloomberg's reporting, cited by Decrypt, Google held Pro back specifically because it fell short of internal quality targets, particularly on coding tasks. That's a more specific explanation than Google has offered publicly, and it lines up with independent testing. Decrypt's own hands-on evaluation of Gemini 3.6 Flash, the model that did ship, found its coding output underwhelming, producing malformed HTML in a simple test that required multiple attempts to get working. If the Flash-tier model that did make it out the door is still struggling with coding reliability, it's not hard to see why the flagship model built to handle harder tasks is taking longer.

Independent benchmarks add more context. On GDPval-AA v2, a knowledge-work benchmark scored on a chess-style Elo rating, Anthropic's Claude Sonnet 5 currently leads at 1607, well ahead of Gemini 3.6 Flash's 1421. That gap isn't a fair comparison between a flagship and a Flash-tier model, but it does underline the competitive pressure sitting behind Google's decision to keep Pro in testing rather than ship something that might land closer to that gap than Google wants.

A Tease That Reads Two Different Ways

In the same announcement window, Google DeepMind's Logan Kilpatrick posted on X that the team has started what he called its most ambitious pre-training run yet, for Gemini 4, the generation after the one that's currently stuck. That's a striking thing to announce publicly while the previous flagship still hasn't shipped.

Read one way, it's confidence: Google is far enough along on Pro internally that the team already has bandwidth to look past it toward Gemini 4. Read the other way, it's a deflection, an attempt to redirect attention toward an exciting future release rather than dwell on a delayed current one. Both readings are plausible, and Google's own messaging doesn't rule either one out.

What the Pattern Actually Says

Strip away the individual model names and a clearer pattern emerges. Google is proving it can ship efficient, cost-competitive models in its Flash tier reliably and on a reasonable cadence, evidenced by three genuine releases in a single day. What it's struggling to do consistently is ship its most capable, flagship-tier model on the timeline it announces publicly. That's not a small distinction for enterprise customers deciding which company to build long-term AI infrastructure around. A company that reliably ships good mid-tier models but repeatedly slips on its top-tier releases is a fundamentally different bet than a company executing cleanly across its entire lineup.

For now, the Flash tier keeps improving on a real cadence, and Pro keeps sliding further into an undefined future. Google is calling that a positive, evidence that quality control is winning out over shipping speed. Competitors watching the gap between Google's announced deadlines and its actual delivery dates are, understandably, reading it as an opening.

ShareWhatsAppTwitterLinkedIn
AB

Written by

Mr. Aayush Bhatt

Software Engineer interested in how models work and where they fail.

โ† Back to AI