Back to glossary

Term

Gemini 3.7 Flash

Gemini 3.7 Flash (August 2026) is Googles fast workhorse model for coding and agents — 1M-token context, introductory price of 0.75/3.75 USD per million tokens, half the cost of 3.6 Flash.

Gemini 3.7 Flash — explained in more detail

Google announced Gemini 3.7 Flash on August 13, 2026 as its most intelligent workhorse model yet for coding and agents. The multimodal model is positioned as the fast, low-cost workhorse of the Gemini range and targets coding, web development, knowledge-based tasks and agent workflows. It arrived only about three weeks after its predecessor 3.6 Flash and keeps that model’s context window of 1,048,576 tokens.

Example / Practical use

The introductory price is 0.75 USD per 1M input tokens and 3.75 USD per 1M output tokens — making Gemini 3.7 Flash roughly half the cost of the predecessor 3.6 Flash. The model is available through the Gemini API in Google AI Studio, in Android Studio, in Google Antigravity for agent-first workflows, and across the Gemini Enterprise Agent Platform and the Gemini app. For teams running many requests in coding or agent pipelines, the low price at high speed is the central advantage.

Distinction

Flash is Googles fast, low-cost line below the more powerful Pro models (such as Gemini 3 Pro). While Pro targets maximum reasoning depth, Flash stands for throughput and low cost per token. Notable at launch: Gemini 3.7 Flash arrived while the top model Gemini 3.5 Pro was still delayed — Google explicitly positioned the Flash model here as a capable everyday workhorse.

See everything in one place:Gemini Flash