Comparison · Updated
Claude Sonnet 5 vs Claude Sonnet 4.6 Cost: How to Compare Without Guessing
Rates move: provider rate cards change without notice, and any figure printed here would rot. This page therefore keeps the method — what actually drives the bill for Claude Sonnet 5 and Claude Sonnet 4.6 — and links to the live rates for both models. Numbers you can reproduce beat numbers you can quote.
The short answer
Mid-tier generation gap: re-rate your measured workload on both rate cards before migrating.
What drives the cost
For both models the bill is the same arithmetic: input tokens multiplied by the input rate, plus output tokens multiplied by the output rate, divided by one million. Three things move the result. The output multiplier: generation is sequential compute, so output is priced above input. The context weight: long contexts are billed as input on every call, and multi-turn workloads re-send history. Cache rules: cached input is usually discounted, and the discount changes the ranking of rate cards that look close on paper.
Claude Sonnet 5: workload shape
Current mid tier. Its cost profile is the rate card applied to that shape; the current values live on the Claude Sonnet 5 model page.
Claude Sonnet 4.6: workload shape
Previous mid tier with broad production use. Same arithmetic, different rate card — see the Claude Sonnet 4.6 model page for the current values.
The comparison with live numbers
Open Claude Sonnet 5 and Claude Sonnet 4.6 for the current per-million rates, then enter the same token counts for both in the LLM cost calculator. Comparing two rate cards on one workload shape is the only comparison that means anything; comparing headline rates on paper is how budgets get surprised.
What this estimate excludes
Taxes, platform fees, cache-specific billing rules beyond the published unit rates, minimum purchases, failed requests and retries. Verify a real bill against the provider's usage records.
Related comparisons
- Claude Opus 5 vs Claude Sonnet 4.6 Cost: How to Compare Without Guessing
Claude Opus 5 vs Claude Sonnet 4.6 cost compared by method: workload shape, output multipliers, context weight and live rates. No stale figures.
- Claude Sonnet 4.6 vs GPT-5.4 Mini Cost: How to Compare Without Guessing
Claude Sonnet 4.6 vs GPT-5.4 Mini cost compared by method: workload shape, output multipliers, context weight and live rates. No stale figures.
- Claude Sonnet 5 vs GPT-5.4 Cost: How to Compare Without Guessing
Claude Sonnet 5 vs GPT-5.4 cost compared by method: workload shape, output multipliers, context weight and live rates. No stale figures.
- Kimi K3 vs Claude Sonnet 5 Cost: How to Compare Without Guessing
Kimi K3 vs Claude Sonnet 5 cost compared by method: workload shape, output multipliers, context weight and live rates. No stale figures.
Estimate the cost
- Claude API cost calculator
Estimate Anthropic API costs from input and output token counts. Formula, output multiplier and monthly projection.
- All LLM cost calculators
Calculators by provider, workload and budget, all built on the same two-rate token formula.
- How to estimate LLM API costs
The token-first method: measure one run, model the workload shape, project monthly, reconcile usage.