Term
Gemini 3.7 Flash
Gemini 3.7 Flash (August 2026) is Googles fast workhorse model for coding and agents — 1M-token context, introductory price of 0.75/3.75 USD per million tokens, half the cost of 3.6 Flash.
Gemini 3.7 Flash — explained in more detail
Google announced Gemini 3.7 Flash on August 13, 2026 as its most intelligent workhorse model yet for coding and agents. The multimodal model is positioned as the fast, low-cost workhorse of the Gemini range and targets coding, web development, knowledge-based tasks and agent workflows. It arrived only about three weeks after its predecessor 3.6 Flash and keeps that model’s context window of 1,048,576 tokens.
Example / Practical use
The introductory price is 0.75 USD per 1M input tokens and 3.75 USD per 1M output tokens — making Gemini 3.7 Flash roughly half the cost of the predecessor 3.6 Flash. The model is available through the Gemini API in Google AI Studio, in Android Studio, in Google Antigravity for agent-first workflows, and across the Gemini Enterprise Agent Platform and the Gemini app. For teams running many requests in coding or agent pipelines, the low price at high speed is the central advantage.
Distinction
Flash is Googles fast, low-cost line below the more powerful Pro models (such as Gemini 3 Pro). While Pro targets maximum reasoning depth, Flash stands for throughput and low cost per token. Notable at launch: Gemini 3.7 Flash arrived while the top model Gemini 3.5 Pro was still delayed — Google explicitly positioned the Flash model here as a capable everyday workhorse.