OpenAI pricing profile

GPT-5.4 Mini

Smaller reasoning model for lower-cost agent and coding tasks.

activebalanced400,000 contextVerified 2026-08-02

Pricing modes

USD per one million tokens. A documented long-context band applies to the full request when its threshold is crossed.

ModeInputCached inputCache writeOutputLong-context band
standard$0.75$0.075$4.5
batch$0.375$0.0375$2.25
fast$1.5$0.15$9
Worked example

100K input + 2K output

$0.0840

Estimated standard-tier token cost for one request. At 10,000 identical requests per month, the token subtotal would be $840.00.

Open this scenario →
Technical limits

Model constraints

Context window400,000 tokens
Maximum output128,000
Long-context thresholdNo separate band recorded
Tokenizer profileopenai

What can change the invoice

  • Cached-token savings depend on provider eligibility and correct cache implementation.
  • Eligible regional processing endpoints can add 10% to token charges.
  • Web search is separately priced at $10 per 1,000 calls in the public list-price model.
  • Taxes, cloud marketplaces, enterprise contracts, retries and quality differences are not included in this example.

Related OpenAI models

Compare models from the same provider before evaluating quality and latency.