Google ships three new Gemini 3.5 Flash models as flagship 3.5 Pro slips

Google's Tuesday drop was all about the workhorse tier. Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber target developers building AI agents at scale, emphasizing token efficiency, speed and reliability rather than frontier capability. Jeff Dean highlighted that 3.6 Flash is 'much more token efficient' than 3.5 Flash — reports cite ~17% lower output token usage — with a side-by-side demo. Flash Cyber, per The Hacker News, is a security-oriented variant, notable given the week's cyber theme.
The subtext is the missing flagship: Gemini 3.5 Pro remains unavailable, still undergoing partner testing amid reports of internal delays and morale strain inside DeepMind. Commentators (Gizmodo) read the Flash release as Google 'reminding you it has an AI model too' while its top-tier answer to GPT-5.6 and Claude Fable 5 slips. Logan Kilpatrick separately said Google has begun its 'most ambitious pre-training run yet' for Gemini 4, suggesting attention may be leapfrogging to the next generation.
Google's momentum elsewhere is real: Sundar Pichai reported 24% YoY Alphabet revenue growth, Google Cloud accelerating to 82% growth, the Gemini app at 950M MAUs, and model APIs processing 22B tokens/min (up from 16B+), driven by exactly these Flash models. Gemini Enterprise is used by 90% of the Fortune 100.
The community verdict was harsh on capability, though: an r/OpenAI thread titled 'Gemini 3.6 Flash: twice as fast, 18% cheaper, and precisely 0% smarter' (389 upvotes) captured skepticism that efficiency gains mask stalled reasoning progress. What to watch: whether 3.5 Pro ships at all before Gemini 4 pre-training bears fruit.