About Gemini 3.7 Flash
Gemini 3.7 Flash is a Flash-tier workhorse of Google's Gemini 3 line, which Google pitches at complex coding, agentic workflows and reliable multi-step execution. Like the rest of the family it is natively multimodal — text, images, audio, video and PDF documents all go in, text comes out — and it carries the line's million-token-class context window with a generous output allowance.
It ships as a stable release rather than a preview; Google's listing places it between Gemini 3.6 Flash and Gemini 3.8 Flash, and all three remain available at matching rates.
The pricing here comes with a calendar attached: Google lists time-limited introductory rates that expire on a stated date, after which input and output both double — the callout above carries the figures and the deadline, and budgeting against the post-rise rates is the safer default for anything long-lived. Context caching discounts repeated input by an order of magnitude but adds an hourly storage fee for keeping the cache alive, so it pays off only above a steady request rate. Batch and flex processing halve the rates again for work that can wait, which stacks with the introductory window while it lasts.
Within Google's lineup the practical question is whether a task needs the newer generation at all: the older Gemini 2.5 Flash and Flash-Lite tiers still list well below this model and remain the budget picks for straightforward high-volume text work, while agentic pipelines and multimodal-heavy jobs are what Google positions this tier for. The pricing page also lists a priority tier at a premium over standard rates, so the same model spans a wide per-token range depending on how urgently you need answers.
Frequently asked questions
How much does the Gemini 3.7 Flash API cost?
Gemini 3.7 Flash costs $0.75 per million input tokens and $3.75 per million output tokens, with cached input at $0.075 per million (90% off). Note: Introductory pricing to 2026-12-31; from 2027-01-01 it is $1.50 in / $7.50 out. Cached input also incurs $0.50/MTok/hour storage.
How much does a typical chat message cost with Gemini 3.7 Flash?
A short chat message (about 500 input and 250 output tokens) costs roughly $0.0013 with Gemini 3.7 Flash. At 1,000 such requests a day, that is about $39.38 per month.
Is Gemini 3.7 Flash cheap compared to similar models?
Gemini 3.7 Flash has the #10 cheapest input rate of the 29 models we track. On a mixed workload (10K input / 2K output), it costs $0.015 per request and $0.020 for o4-mini.
What does Gemini 3.7 Flash cost per 1,000 tokens?
Per 1,000 tokens, Gemini 3.7 Flash costs $0.00075 for input and $0.0037 for output, or $0.00007 for cached input. Providers quote rates per million tokens ($0.75/M in, $3.75/M out here), so divide by a thousand for the per-1K figure older pricing pages used.
How large a prompt can Gemini 3.7 Flash take?
Gemini 3.7 Flash accepts up to 1024K tokens of input context and can generate up to 64K tokens of output per request. At its input rate, filling the entire 1024K-token window costs about $0.786 per request before any output.
Related
Compare all models at once ·Count tokens in your prompt
All Google model costs: Gemini 3.8 Flash · Gemini 3.6 Flash · Gemini 3.1 Pro Preview · Gemini 2.5 Pro · Gemini 2.5 Flash · Gemini 2.5 Flash-Lite
Last updated 2026-09-03. Prices verified against Google's official pricing page; see the methodology.