Worked example
100K input + 2K output
$0.0224
Estimated standard-tier token cost for one request. At 10,000 identical requests per month, the token subtotal would be $224.00.
Open this scenario →Cost-sensitive GPT-5.6 model for high-volume workloads.
USD per one million tokens. A documented long-context band applies to the full request when its threshold is crossed.
| Mode | Input | Cached input | Cache write | Output | Long-context band |
|---|---|---|---|---|---|
| standard | $0.2 | $0.02 | $0.25 | $1.2 | Input $0.4 / output $1.8 |
| batch | $0.1 | $0.01 | $0.125 | $0.6 | Input $0.2 / output $0.9 |
| flex | $0.1 | $0.01 | $0.125 | $0.6 | Input $0.2 / output $0.9 |
| fast | $0.4 | $0.04 | $0.5 | $2.4 | — |
Estimated standard-tier token cost for one request. At 10,000 identical requests per month, the token subtotal would be $224.00.
Open this scenario →Pricing was checked against first-party documentation on 2026-08-02. The active schedule is effective from 2026-07-30.
Compare models from the same provider before evaluating quality and latency.
Frontier reasoning and coding model for complex professional work.
View pricing →Balances intelligence and cost for production applications.
View pricing →Smaller reasoning model for lower-cost agent and coding tasks.
View pricing →