Gemini 3.5 Flash Launch: Benchmarks, Speed, Pro Still Pending

Google released Gemini 3.5 Flash before Pro, claiming 76.2% on Terminal-Bench 2.1 and faster agentic work while promising Pro for the following month.

Flash shipped. Pro was a promise.

Google launched Gemini 3.5 Flash on May 19, 2026, across the Gemini app, AI Mode in Search, the Gemini API through AI Studio, Android Studio, its Antigravity development platform, and the Gemini Enterprise Agent Platform. Gemini 3.5 Pro, the larger model in the same generation, was not part of the release. Google said Pro was already in internal use and would follow the next month. [1]

Atlas interpretation: Frontier labs have usually led with the largest model and backfilled smaller, cheaper variants afterward. Shipping the small model first and holding the flagship back reverses that order, and it is worth treating as a deliberate choice rather than a scheduling accident: Google had a specific product argument for why Flash could go out alone. [1][3]

The pitch was Flash as the workhorse, not the placeholder

Google framed the two models as a pair rather than a hierarchy. Tulsee Doshi, the company's senior director and head of product for Gemini, described the intended split to reporters: "3.5 Pro becomes your orchestrator, your planner, and then it actually can leverage Flash to be the various sub-agents." On that framing, Flash was never meant to be a stopgap; it is the model built to run in volume underneath a Pro that plans the work. [3]

DeepMind's chief technologist, Koray Kavukcuoglu, told the same outlet that Flash "outperforms our latest frontier model, 3.1 Pro, on nearly all the benchmarks," and that it ran "4x faster than other frontier models," with an optimized configuration reaching "12x faster with the same quality." Those are the company's own comparisons; the article carries no independent benchmark run or third-party confirmation of the speed multiples. [3]

The numbers, and what they don't establish

Gemini 3.5 Flash's model card lists 76.2% on Terminal-Bench 2.1, an agentic terminal-coding benchmark; 1656 Elo on GDPval-AA, a knowledge-work comparison; 83.6% on MCP Atlas, a multi-step tool-use benchmark; and 84.2% on CharXiv Reasoning, a chart-and-figure interpretation task. The card states these evaluations are automated, not human-reviewed or red-teamed, and warns that "performance results reported below are computed with improved evaluations and thus are not directly comparable with performance results found in previous Gemini model cards." [2]

Atlas interpretation: That caveat cuts both ways. It means the 76.2% figure cannot be read against an older Gemini card's Terminal-Bench number as a clean before-and-after, and it is Google's own acknowledgment that the yardstick moved along with the model. The Terminal-Bench leaderboard itself has since moved past version 2.1 to a newer revision, which is the same pattern seen elsewhere on this timeline: a benchmark score is a statement about a specific test at a specific moment, not a fixed unit that keeps its meaning as the underlying suite is revised. [2][4]

One release inside a fast-moving generation

Gemini 3.5 Flash arrived three months after Gemini 3.1 Pro, which itself had followed the original Gemini 3 launch by three months. Comparing 3.5 Flash against 3.1 Pro, as Google's own announcement does, is a comparison across that same generation rather than against a competitor's model. [1]

Atlas interpretation: Read against that cadence, the Flash-first release looks less like a one-off reordering and more like Google using its fastest-moving tier to keep a model shipping every few months while the larger Pro model gets more internal runway before its own release. [1][3]

Sources

  1. Gemini 3.5: frontier intelligence with action

    Google · May 19, 2026

  2. Gemini 3.5 Flash Model Card

    Google DeepMind · May 19, 2026

  3. With Gemini 3.5 Flash, Google bets its next AI wave on agents, not chatbots

    TechCrunch · May 19, 2026

  4. Terminal-Bench Leaderboard

    Terminal-Bench (Stanford, Harbor, and the Laude Institute) · Sep 8, 2026