Gemini 3.6 Flash API Cost Calculator

Google's Gemini 3.6 Flash: $0.75/M input · $3.75/M output · $0.075/M cached input. Price your own workload below.

Introductory pricing to 2026-12-31; from 2027-01-01 it is $1.50 in / $7.50 out. Cached input also incurs $0.50/MTok/hour storage.

Presets:

What Gemini 3.6 Flash costs in practice

WorkloadInputOutputCost
Short chat message500250$0.0013
10-page document + summary7,000500$0.0071
Agent / coding session50,00010,000$0.075
1M tokens in + 1M out1,000,0001,000,000$4.50

Gemini 3.6 Flash vs nearest rivals

Closest-priced alternatives on a mixed 10K-input / 2K-output request:

ModelInput $/MTokOutput $/MTokMixed request
Gemini 3.6 Flash$0.75$3.75$0.015
o4-mini$1.1$4.4$0.020
Claude Haiku 4.5$1$5$0.020
DeepSeek V4 Pro$1.32$3.96$0.021

About Gemini 3.6 Flash

Gemini 3.6 Flash occupies the Flash slot in Google's Gemini 3 line: the tier Google describes as balancing speed and multimodal capability across general agentic and everyday tasks. Like the rest of the family it is natively multimodal well beyond text and images — audio and video input are part of the Gemini design — and it carries the line's million-token-class context window.

It is a stable release rather than a preview, and Google's model listing places a newer Flash generation above it, the usual rhythm of a line that iterates quickly while keeping prior stable models available.

A generational quirk defines its economics: Gemini 3.6 Flash lists above the older Gemini 2.5 Flash while carrying the newer generation's capabilities, so 'Flash' no longer means what it cost a generation ago — check whether your workload actually needs the newer line before assuming the name implies budget pricing. Google prices it in lockstep with Gemini 3.7 Flash, including the time-limited introductory rates that later step up (the callout above has the figures), so there is no price reason to prefer the older of the two. Context caching discounts repeated input by an order of magnitude but adds an hourly storage fee for keeping the cache alive, so caching pays off only above a steady request rate. Batch processing brings a further discount for asynchronous work.

Frequently asked questions

How much does the Gemini 3.6 Flash API cost?

Gemini 3.6 Flash costs $0.75 per million input tokens and $3.75 per million output tokens, with cached input at $0.075 per million (90% off). Note: Introductory pricing to 2026-12-31; from 2027-01-01 it is $1.50 in / $7.50 out. Cached input also incurs $0.50/MTok/hour storage.

How much does a typical chat message cost with Gemini 3.6 Flash?

A short chat message (about 500 input and 250 output tokens) costs roughly $0.0013 with Gemini 3.6 Flash. At 1,000 such requests a day, that is about $39.38 per month.

Is Gemini 3.6 Flash cheap compared to similar models?

Gemini 3.6 Flash has the #10 cheapest input rate of the 29 models we track. On a mixed workload (10K input / 2K output), it costs $0.015 per request and $0.020 for o4-mini.

What does Gemini 3.6 Flash cost per 1,000 tokens?

Per 1,000 tokens, Gemini 3.6 Flash costs $0.00075 for input and $0.0037 for output, or $0.00007 for cached input. Providers quote rates per million tokens ($0.75/M in, $3.75/M out here), so divide by a thousand for the per-1K figure older pricing pages used.

How large a prompt can Gemini 3.6 Flash take?

Gemini 3.6 Flash accepts up to 1024K tokens of input context and can generate up to 64K tokens of output per request. At its input rate, filling the entire 1024K-token window costs about $0.786 per request before any output.

Related

Compare all models at once ·Count tokens in your prompt

All Google model costs: Gemini 3.8 Flash · Gemini 3.7 Flash · Gemini 3.1 Pro Preview · Gemini 2.5 Pro · Gemini 2.5 Flash · Gemini 2.5 Flash-Lite

Last updated 2026-09-03. Prices verified against Google's official pricing page; see the methodology.