About Gemini 2.5 Pro
Gemini 2.5 Pro was the model that established Google's thinking-model approach in general availability — Google describes it as the most advanced of the Gemini 2.5 family, built for complex tasks with deep reasoning and coding capability. It remains a stable, fully supported release with the family's signature range: multimodal input including audio and video, and a million-token-class context window.
With the Gemini 3 line above it, its role has shifted from frontier to known quantity — the version teams stay on when validated behaviour matters more than headline capability.
The list rates undercut the newer Pro preview, but the same two-tier structure applies: prompts past the long-context threshold bill at elevated rates for the whole request (see the callout above), and cached input carries the Pro-level hourly storage fee on top of its discounted token rate. That combination rewards a specific shape of workload — moderate prompt sizes at steady volume — and punishes occasional very-long-context use. If you are choosing between Pro generations, the question is capability per dollar on your actual task; the mixed-workload comparison above is the place to start.
Frequently asked questions
How much does the Gemini 2.5 Pro API cost?
Gemini 2.5 Pro costs $1.25 per million input tokens and $10 per million output tokens, with cached input at $0.125 per million (90% off). Note: Prompts over 200K tokens: $2.50 in / $15 out. Cache storage $4.50/MTok/hour.
How much does a typical chat message cost with Gemini 2.5 Pro?
A short chat message (about 500 input and 250 output tokens) costs roughly $0.0031 with Gemini 2.5 Pro. At 1,000 such requests a day, that is about $93.75 per month.
Is Gemini 2.5 Pro cheap compared to similar models?
Gemini 2.5 Pro has the #15 cheapest input rate of the 29 models we track. On a mixed workload (10K input / 2K output), it costs $0.033 per request, versus $0.030 for Mistral Medium 3.5 and $0.036 for o3.
What does Gemini 2.5 Pro cost per 1,000 tokens?
Per 1,000 tokens, Gemini 2.5 Pro costs $0.0013 for input and $0.0100 for output, or $0.00013 for cached input. Providers quote rates per million tokens ($1.25/M in, $10/M out here), so divide by a thousand for the per-1K figure older pricing pages used.
How large a prompt can Gemini 2.5 Pro take?
Gemini 2.5 Pro accepts up to 1024K tokens of input context and can generate up to 64K tokens of output per request. At its input rate, filling the entire 1024K-token window costs about $1.31 per request before any output.
Related
Compare all models at once ·Count tokens in your prompt · Gemini 2.5 Pro vs GPT-4o head-to-head
All Google model costs: Gemini 3.8 Flash · Gemini 3.7 Flash · Gemini 3.6 Flash · Gemini 3.1 Pro Preview · Gemini 2.5 Flash · Gemini 2.5 Flash-Lite
Last updated 2026-09-03. Prices verified against Google's official pricing page; see the methodology.