Gemini 2.5 Flash API Cost Calculator

Google's Gemini 2.5 Flash: $0.3/M input · $2.5/M output · $0.03/M cached input. Price your own workload below.

Cached input also incurs $1.00/MTok/hour storage.

Presets:

What Gemini 2.5 Flash costs in practice

WorkloadInputOutputCost
Short chat message500250$0.0008
10-page document + summary7,000500$0.0034
Agent / coding session50,00010,000$0.040
1M tokens in + 1M out1,000,0001,000,000$2.80

Gemini 2.5 Flash vs nearest rivals

Closest-priced alternatives on a mixed 10K-input / 2K-output request:

ModelInput $/MTokOutput $/MTokMixed request
Gemini 2.5 Flash$0.3$2.5$0.0080
DeepSeek V4 Pro$0.435$0.87$0.0061
GPT-5 mini$0.25$2$0.0065
Mistral Large 3$0.5$1.5$0.0080

Frequently asked questions

How much does the Gemini 2.5 Flash API cost?

Gemini 2.5 Flash costs $0.3 per million input tokens and $2.5 per million output tokens, with cached input at $0.03 per million (90% off). Note: Cached input also incurs $1.00/MTok/hour storage.

How much does a typical chat message cost with Gemini 2.5 Flash?

A short chat message (about 500 input and 250 output tokens) costs roughly $0.0008 with Gemini 2.5 Flash. At 1,000 such requests a day, that is about $23.25 per month.

Is Gemini 2.5 Flash cheap compared to similar models?

Gemini 2.5 Flash has the #8 cheapest input rate of the 26 models we track. On a mixed workload (10K input / 2K output), it costs $0.0080 per request, versus $0.0061 for DeepSeek V4 Pro.

Related

Compare all models at once ·Count tokens in your prompt

Last updated 2026-08-03. Prices verified againstGoogle's official pricing page; see themethodology.