Skip to content

Google

Gemini 2.5 Pro: price, context window, and caching

The previous-generation Pro, still stable, and still Gemini CLI's Pro fallback for accounts without preview access.

Released June 17, 2025 · Prices as of September 28, 2026

In Google’s words

“A Pro model which excels at coding and complex reasoning tasks.”

Google: Gemini API pricing

What Google says it’s good at

  • Reasoning over complex problems in code, math, and STEM Source
  • Analyzing large codebases and documents with long context Source

Facts

Specs and prices

FactGemini 2.5 Pro
MakerGoogle
API model idgemini-2.5-pro
ReleasedJune 17, 2025
StatusPrevious generation
Context window1.05M tokens
Max output65.5K tokens
Open weightsNo
Input, per 1M tokens$1.25
Cache hit, per 1M$0.125
Cache write, per 1M$1.25 (same as input)
Output, per 1M tokens$10
Runs inGemini CLI, OpenCode, and OpenRouter

Standard API rates in US dollars, as published by the maker on September 28, 2026. Batch and priority tiers, taxes, and subscription plans are not included. Gemini 2.5 Pro: Prompts over 200K input tokens cost $2.50 input, $0.25 cached, and $15 output per million tokens.

Good to know

  • Since September 18, 2026, Google limits access to accounts that used it before. It is not deprecated.
  • It uses thinking budgets rather than thinking levels, and Gemini CLI sends a budget of 8,192 tokens.

Cost

What typical work costs

Example token counts at Gemini 2.5 Pro’s published rates. On the agentic session, caching saves $2.25 against billing every token as ordinary input.

Example workload costs for Gemini 2.5 Pro
Agentic coding session, 100K input, 400K written to cache, 2M read from cache, 50K output$1.38
Large one-off review, 150K input with no cache hits, 10K output$0.29
Output-heavy generation, 30K input, 80K output$0.84
A month of sessions, 110 sessions: 5 a day, 22 working days$151.25

Prompt caching

How Google bills cached tokens

Google

Implicit caching is on by default for Gemini 2.5 and newer. When a request repeats a prefix Google has cached, the discount is applied automatically, but a hit is not guaranteed.

Explicit caching creates a cache you reference by name, with a guaranteed discount. It adds a storage charge for as long as the cache lives, 1 hour by default: $4.50 per million tokens per hour on Pro models and $0.50 to $1 on Flash models.

On every model compared here, a cache hit costs 10% of the input price. Google publishes no separate cache-write price, so this blog prices written tokens as ordinary input.

Source: Google: Context caching

Gemini 2.5 Pro compared

  • Claude Sonnet 4.6 vs Gemini 2.5 Pro

    Gemini 2.5 Pro costs less than Claude Sonnet 4.6 on every rate, but caps output at 65.5K tokens and now limits who can use it. The full cost and access picture.

  • Gemini 2.5 Pro vs Gemini 2.5 Flash

    Gemini 2.5 Pro costs about 4x Gemini 2.5 Flash, and since September 18, 2026, only earlier users can reach either. What each costs and where to go next.

  • Gemini 3.1 Pro Preview vs Gemini 2.5 Pro

    Gemini 3.1 Pro Preview costs 60% more per input token than Gemini 2.5 Pro and is still a preview, while 2.5 Pro now limits new access. How to choose.

Your own numbers

See what Gemini 2.5 Pro really costs you.

everyaitoken reads your Gemini CLI, OpenCode, and OpenRouter history on your Mac and prices every request at API rates, with what caching saved or cost. $9 once.

Launching soonSee the cache math