Gemini 2.5 Flash API Cost Calculator

Google's Gemini 2.5 Flash: $0.3/M input · $2.5/M output · $0.03/M cached input. Price your own workload below.

Cached input also incurs $1.00/MTok/hour storage.

Presets:

What Gemini 2.5 Flash costs in practice

WorkloadInputOutputCost
Short chat message500250$0.0008
10-page document + summary7,000500$0.0034
Agent / coding session50,00010,000$0.040
1M tokens in + 1M out1,000,0001,000,000$2.80

Gemini 2.5 Flash vs nearest rivals

Closest-priced alternatives on a mixed 10K-input / 2K-output request:

ModelInput $/MTokOutput $/MTokMixed request
Gemini 2.5 Flash$0.3$2.5$0.0080
GPT-5 mini$0.25$2$0.0065
DeepSeek V4 Flash$0.44$1.32$0.0070
Mistral Large 3$0.5$1.5$0.0080

About Gemini 2.5 Flash

Gemini 2.5 Flash is Google's stated price-performance pick within the Gemini 2.5 family: a low-latency, high-volume model that still supports reasoning, rather than a stripped-down budget option. It keeps the family's full multimodal intake — text, images, audio, video — and the million-token-class context window, which is unusual generosity at its price tier.

Google has since shipped Flash models in newer generations, but this release stays available as a stable workhorse, and its combination of low rates and real reasoning keeps it a common default for production pipelines.

The economics are the argument: rates an order of magnitude below the newer Flash generation, with cached input discounted a further order below that. The hourly cache-storage fee (callout above) is the fine print — at low traffic the storage cost can outweigh the token savings, so caching earns its keep only once requests arrive steadily. Output is priced at several times input, standard for the industry, so response length still dominates chatty workloads. Flash-Lite sits below for pure-throughput tasks; the step up to a Pro tier is for depth, not speed.

Frequently asked questions

How much does the Gemini 2.5 Flash API cost?

Gemini 2.5 Flash costs $0.3 per million input tokens and $2.5 per million output tokens, with cached input at $0.03 per million (90% off). Note: Cached input also incurs $1.00/MTok/hour storage.

How much does a typical chat message cost with Gemini 2.5 Flash?

A short chat message (about 500 input and 250 output tokens) costs roughly $0.0008 with Gemini 2.5 Flash. At 1,000 such requests a day, that is about $23.25 per month.

Is Gemini 2.5 Flash cheap compared to similar models?

Gemini 2.5 Flash has the #7 cheapest input rate of the 29 models we track. On a mixed workload (10K input / 2K output), it costs $0.0080 per request, versus $0.0065 for GPT-5 mini.

What does Gemini 2.5 Flash cost per 1,000 tokens?

Per 1,000 tokens, Gemini 2.5 Flash costs $0.00030 for input and $0.0025 for output, or $0.00003 for cached input. Providers quote rates per million tokens ($0.3/M in, $2.5/M out here), so divide by a thousand for the per-1K figure older pricing pages used.

How large a prompt can Gemini 2.5 Flash take?

Gemini 2.5 Flash accepts up to 1024K tokens of input context and can generate up to 64K tokens of output per request. At its input rate, filling the entire 1024K-token window costs about $0.315 per request before any output.

Related

Compare all models at once ·Count tokens in your prompt · Gemini 2.5 Flash vs GPT-5 mini head-to-head

All Google model costs: Gemini 3.8 Flash · Gemini 3.7 Flash · Gemini 3.6 Flash · Gemini 3.1 Pro Preview · Gemini 2.5 Pro · Gemini 2.5 Flash-Lite

Last updated 2026-09-03. Prices verified against Google's official pricing page; see the methodology.