OpenAI pricing profile

GPT-5.6 Luna

Cost-sensitive GPT-5.6 model for high-volume workloads.

activeeconomy1,050,000 contextVerified 2026-08-02

Pricing modes

USD per one million tokens. A documented long-context band applies to the full request when its threshold is crossed.

ModeInputCached inputCache writeOutputLong-context band
standard$0.2$0.02$0.25$1.2Input $0.4 / output $1.8
batch$0.1$0.01$0.125$0.6Input $0.2 / output $0.9
flex$0.1$0.01$0.125$0.6Input $0.2 / output $0.9
fast$0.4$0.04$0.5$2.4
Worked example

100K input + 2K output

$0.0224

Estimated standard-tier token cost for one request. At 10,000 identical requests per month, the token subtotal would be $224.00.

Open this scenario →
Technical limits

Model constraints

Context window1,050,000 tokens
Maximum output128,000
Long-context threshold272,000
Tokenizer profileopenai

What can change the invoice

  • Requests above 272,000 input tokens can move into a higher long-context price band.
  • Cached-token savings depend on provider eligibility and correct cache implementation.
  • Eligible regional processing endpoints can add 10% to token charges.
  • Web search is separately priced at $10 per 1,000 calls in the public list-price model.
  • Taxes, cloud marketplaces, enterprise contracts, retries and quality differences are not included in this example.

Related OpenAI models

Compare models from the same provider before evaluating quality and latency.