Google pricing profile

Gemini 2.5 Pro

Long-context Gemini reasoning model with a 200K pricing threshold.

activefrontier1,000,000 contextVerified 2026-08-02

Pricing modes

USD per one million tokens. A documented long-context band applies to the full request when its threshold is crossed.

ModeInputCached inputCache writeOutputLong-context band
standard$1.25$0.125$10Input $2.5 / output $15
batch$0.625$0.125$5Input $1.25 / output $7.5
flex$0.625$0.125$5Input $1.25 / output $7.5
priority$2.25$0.225$18Input $4.5 / output $27
Worked example

100K input + 2K output

$0.1450

Estimated standard-tier token cost for one request. At 10,000 identical requests per month, the token subtotal would be $1,450.00.

Open this scenario →
Technical limits

Model constraints

Context window1,000,000 tokens
Maximum output65,536
Long-context threshold200,000
Tokenizer profilegoogle

What can change the invoice

  • Requests above 200,000 input tokens can move into a higher long-context price band.
  • Cached-token savings depend on provider eligibility and correct cache implementation.
  • Web search is separately priced at $35 per 1,000 calls in the public list-price model.
  • Taxes, cloud marketplaces, enterprise contracts, retries and quality differences are not included in this example.

Related Google models

Compare models from the same provider before evaluating quality and latency.