OpenAI pricing profile

GPT-5.6 Terra

Balances intelligence and cost for production applications.

activebalanced1,050,000 contextVerified 2026-08-02

Pricing modes

USD per one million tokens. A documented long-context band applies to the full request when its threshold is crossed.

ModeInputCached inputCache writeOutputLong-context band
standard$2$0.2$2.5$12Input $4 / output $18
batch$1$0.1$1.25$6Input $2 / output $9
flex$1$0.1$1.25$6Input $2 / output $9
fast$4$0.4$5$24
Worked example

100K input + 2K output

$0.2240

Estimated standard-tier token cost for one request. At 10,000 identical requests per month, the token subtotal would be $2,240.00.

Open this scenario →
Technical limits

Model constraints

Context window1,050,000 tokens
Maximum output128,000
Long-context threshold272,000
Tokenizer profileopenai

What can change the invoice

  • Requests above 272,000 input tokens can move into a higher long-context price band.
  • Cached-token savings depend on provider eligibility and correct cache implementation.
  • Eligible regional processing endpoints can add 10% to token charges.
  • Web search is separately priced at $10 per 1,000 calls in the public list-price model.
  • Taxes, cloud marketplaces, enterprise contracts, retries and quality differences are not included in this example.

Related OpenAI models

Compare models from the same provider before evaluating quality and latency.