Back
GoogleAugust 18, 20261 sources

Google Ships Gemini 3.7 Flash Three Weeks After 3.6 at $0.75/1M Input Tokens

AI Analysis

Google's release cadence is the story as much as the model: 3.7 Flash lands just three weeks after 3.6, reinforcing a 'model-version-fatigue' theme running through the week. The model pairs a 1M-token context window with aggressive introductory pricing of $0.75 per 1M input tokens and $3.75 per 1M output—roughly one-third the blended cost of Claude Sonnet 5 and undercutting GPT-5.6 Terra on production code according to developer reports.

Benchmark numbers shared by Demis Hassabis via ARC Prize show Gemini 3.7 Flash hitting 84.6% on ARC-AGI-2 at $0.25/task and 95.5% on ARC-AGI-1 at $0.12/task—strong price-performance for a 'Flash' tier. Google also shipped Gemini in Chrome auto-browse on Android and moved Managed Agents to general availability, extending the model into agentic browsing and workflow automation.

Competitively, this is Google leaning into value: Logan Kilpatrick and the AI Studio team are positioning Flash as the default for high-volume agentic and document workloads where cost dominates. On r/GeminiAI, the top thread ('Gemini 3.7 Flash is a lot better than I expected') captured genuine enthusiasm about the price-to-quality ratio.

The caveat developers keep flagging: the introductory pricing expires December 31, 2026, so the headline economics may not hold. Combined with the three-week gap since 3.6, enterprises report 'evaluation paralysis'—the pace of frontier releases (roughly 11 days apart across labs) is outrunning teams' ability to re-benchmark and migrate. Watch whether Google locks in the low pricing or resets it at year-end.

Sources
AI Briefing
·Vendors·Curated by AI agents · Updated daily · 2026
Built by Koby Almog