| Input per 1M tokens | $0.75 |
| Output per 1M tokens | $4.5 |
| Cache read per 1M tokens | — |
| Status | active |
| Source | official pricing page · verified 2026-09-07 |
Migration target for gpt-3.5-turbo and gpt-5-mini.
Example costs
- One request (2,000 in / 500 out): $0.0037
- 100K such requests/month: $375
- 1M input + 1M output tokens: $5.25
Compare against other models →
How GPT-5.4 mini compares on rate
GPT-5.4 mini runs $0.75 input / $4.5 output per MTok — a combined $5.25 for 1M tokens each way. The cheapest active model in our table, Gemini 2.5 Flash-Lite at $0.1/$0.4 (combined $0.5), is about 90% cheaper on combined rate (verified 2026-09-07).
Cheapest active OpenAI alternative: GPT-5.6 Luna at $0.2/$1.2 per MTok — combined $1.4 against $5.25, about 73% less.
Per-token rate is not per-task cost — tokenizers differ between model families, so benchmark with your own payloads before switching.
What real volumes cost at GPT-5.4 mini rates
Straight rate-card math at the verified $0.75/$4.5 per-MTok prices, assuming a 75% input / 25% output token split — no caching or batch discounts applied.
| Volume | Split | Cost at GPT-5.4 mini rates |
|---|---|---|
| 100K tokens in a day | 75,000 in + 25,000 out | $0.1688 |
| That pace over a 30-day month (3M tokens) | 2.25M in + 750K out | $5.06 |
| 1M tokens in a month | 750,000 in + 250,000 out | $1.69 |
| Heavy month — 12M tokens | 9M in + 3M out | $20.25 |
Change the split or volume in the calculator — same verified rates, your numbers. Method + the corrections that make estimates real: how to estimate LLM API costs.
Frequently asked questions
What does GPT-5.4 mini cost per million tokens?
GPT-5.4 mini costs $0.75 per 1M input tokens and $4.5 per 1M output tokens on the OpenAI API (verified 2026-09-07).
How much does 100K tokens a day cost on GPT-5.4 mini?
At a 75% input / 25% output split, 100K tokens costs about $0.1688 at GPT-5.4 mini rates — about $5.06 over a 30-day month. That is straight rate-card math, before any caching or batch discounts.
Is there a cheaper model than GPT-5.4 mini?
Yes — the cheapest active model we track is Gemini 2.5 Flash-Lite at $0.1/$0.4 per MTok, a combined $0.5 against GPT-5.4 mini's $5.25 — about 90% less on combined rate (verified 2026-09-07). Tokenizers differ between families, so benchmark per-task cost with your own payloads.
How much do 1M input + 1M output tokens cost on GPT-5.4 mini?
$5.25 — $0.75 for the input million plus $4.5 for the output million, at OpenAI rates verified 2026-09-07.