Worked example
100K input + 2K output
$0.1450
Estimated standard-tier token cost for one request. At 10,000 identical requests per month, the token subtotal would be $1,450.00.
Open this scenario →Long-context Gemini reasoning model with a 200K pricing threshold.
USD per one million tokens. A documented long-context band applies to the full request when its threshold is crossed.
| Mode | Input | Cached input | Cache write | Output | Long-context band |
|---|---|---|---|---|---|
| standard | $1.25 | $0.125 | — | $10 | Input $2.5 / output $15 |
| batch | $0.625 | $0.125 | — | $5 | Input $1.25 / output $7.5 |
| flex | $0.625 | $0.125 | — | $5 | Input $1.25 / output $7.5 |
| priority | $2.25 | $0.225 | — | $18 | Input $4.5 / output $27 |
Estimated standard-tier token cost for one request. At 10,000 identical requests per month, the token subtotal would be $1,450.00.
Open this scenario →Pricing was checked against first-party documentation on 2026-08-02. The active schedule is effective from 2025-01-01.
Compare models from the same provider before evaluating quality and latency.
Fast multimodal Gemini model with search grounding and multiple service tiers.
View pricing →High-volume Gemini model optimized for simple agentic tasks and data processing.
View pricing →Hybrid reasoning model with one-million-token context and low standard pricing.
View pricing →