| Input per 1M tokens | $3 |
| Output per 1M tokens | $15 |
| Cache read per 1M tokens | $0.3 |
| Status | active |
| Source | official pricing page · verified 2026-09-07 |
Example costs
- One request (2,000 in / 500 out): $0.0135
- 100K such requests/month: $1,350
- 1M input + 1M output tokens: $18
Compare against other models →
How Claude Sonnet 4.6 compares on rate
Claude Sonnet 4.6 runs $3 input / $15 output per MTok — a combined $18 for 1M tokens each way. The cheapest active model in our table, Gemini 2.5 Flash-Lite at $0.1/$0.4 (combined $0.5), is about 97% cheaper on combined rate (verified 2026-09-07).
Cheapest active Anthropic alternative: Claude Haiku 4.5 at $1/$5 per MTok — combined $6 against $18, about 67% less.
Cache reads on Claude Sonnet 4.6 are priced at $0.3 per MTok — 0.1× the fresh input rate of $3. A workload where most requests share a long prefix pays far less than the base-rate estimate.
Per-token rate is not per-task cost — tokenizers differ between model families, so benchmark with your own payloads before switching.
What real volumes cost at Claude Sonnet 4.6 rates
Straight rate-card math at the verified $3/$15 per-MTok prices, assuming a 75% input / 25% output token split — no caching or batch discounts applied.
| Volume | Split | Cost at Claude Sonnet 4.6 rates |
|---|---|---|
| 100K tokens in a day | 75,000 in + 25,000 out | $0.6000 |
| That pace over a 30-day month (3M tokens) | 2.25M in + 750K out | $18.00 |
| 1M tokens in a month | 750,000 in + 250,000 out | $6.00 |
| Heavy month — 12M tokens | 9M in + 3M out | $72.00 |
Change the split or volume in the calculator — same verified rates, your numbers. Method + the corrections that make estimates real: how to estimate LLM API costs.
Frequently asked questions
What does Claude Sonnet 4.6 cost per million tokens?
Claude Sonnet 4.6 costs $3 per 1M input tokens and $15 per 1M output tokens on the Anthropic API, with cache reads at $0.3 per 1M tokens (verified 2026-09-07).
How much does 100K tokens a day cost on Claude Sonnet 4.6?
At a 75% input / 25% output split, 100K tokens costs about $0.6000 at Claude Sonnet 4.6 rates — about $18.00 over a 30-day month. That is straight rate-card math, before any caching or batch discounts.
Is there a cheaper model than Claude Sonnet 4.6?
Yes — the cheapest active model we track is Gemini 2.5 Flash-Lite at $0.1/$0.4 per MTok, a combined $0.5 against Claude Sonnet 4.6's $18 — about 97% less on combined rate (verified 2026-09-07). Tokenizers differ between families, so benchmark per-task cost with your own payloads.
How much do cache reads cost on Claude Sonnet 4.6?
Cached input on Claude Sonnet 4.6 is billed at $0.3 per 1M tokens — 0.1× the fresh input rate of $3. Workloads where most requests share a long prefix pay far below the base-rate estimate.