Claude Sonnet 5 Pricing 2026: Real Rates & Hidden Costs
Current Claude Sonnet 5 pricing, the August 31 rate change, Fable 5’s move to credit billing, and where GPT-5.6 and Gemini 3.5 land in the current price war — with the exact math for your own usage.
Quick Answer
Current Claude Sonnet 5 pricing is $2 per million input tokens and $10 per million output tokens, in effect through August 31, 2026 (roughly seven weeks from now).
After that date, Claude Sonnet 5 pricing rises to standard rates of $3/$15 per million tokens — the same rate Sonnet 4.6 already charged.
The catch: Anthropic confirms Sonnet 5’s new tokenizer produces approximately 30% more tokens for the same text, so a workload with an unchanged rate card can still cost more.
Prompt caching (up to 90% savings) and batch processing (up to 50% savings) apply at full value across the entire 1M-token context window.
Third-party testing from Artificial Analysis found that at high reasoning-effort settings, Sonnet 5 can even cost more per task than Opus 4.8 once introductory pricing ends. Full rate table, cache pricing, and the real budgeting math are below.
Calculating days remaining… until Claude Sonnet 5’s introductory pricing ends and standard $3/$15 rates apply on August 31, 2026.
How to Use This Page
Check what changed
Three separate pricing events, covered below in the order they happened.
Scan the live comparison table
Every current model’s rate, side by side, sourced from official pricing pages.
Watch the August 31 cliff
Budget at standard pricing now, not just the introductory rate.
Run your own numbers
Use the AI Token Calculator for your specific volume and model mix.
What Changed in Claude Sonnet 5 Pricing This Month
Three distinct pricing events landed inside two weeks of each other. Most coverage treats them as separate stories.
They’re not. Together they’re a genuine repricing of how teams should budget AI spend for the second half of 2026.
Anthropic released Claude Sonnet 5 on June 30, 2026 and made it the default model for every Free and Pro user starting July 1. Current Claude Sonnet 5 pricing is an introductory $2 per million input tokens and $10 per million output tokens.
That’s a genuine discount off Sonnet 4.6’s standard rate, but it has an expiration date. On August 31, 2026, the rate steps up to $3/$15 — identical to what Sonnet 4.6 already cost.
A week later, on July 8, 2026, Claude Fable 5 moved to pure usage-credit billing at $10 per million input tokens and $50 per million output tokens.
It’s no longer included in any subscription tier; every Pro, Max, Team, and Business user now draws down purchased credits to use it.
Meanwhile OpenAI’s GPT-5.6 entered preview with three tiers — Sol, Terra, and Luna — deliberately spanning a wide price range. Luna undercuts Sonnet 5’s standard rate; Sol prices above it at every level.
The Hidden Part of Claude Sonnet 5 Pricing: Tokens, Not Just Rates
Here’s the detail most launch-day coverage skipped. Sonnet 5’s standard per-token price is unchanged from Sonnet 4.6, at $3/$15.
But its tokenizer segments the same text into finer-grained units. Anthropic’s own Sonnet model documentation confirms the new tokenizer produces approximately 30% more tokens for identical text, with independent testing measuring as high as 35% on code and structured data.
That’s the real story behind this month’s rate card: it looks flat, but still bills more, because the unit being counted changed underneath it.
A workload that looks cost-neutral this week can land 20–30% above your old baseline once the September rate step-up and the tokenizer inflation are counted together.
The practical takeaway: if you’re migrating a production workload during the introductory window, budget Claude Sonnet 5 pricing against the September 1 standard rate and Anthropic’s confirmed ~30% tokenizer inflation — not the July sticker price. The discount is real, but it’s temporary.
How Caching and Batch Processing Change Claude Sonnet 5 Pricing
The per-token rate is only part of what determines an actual Sonnet 5 bill. Three modifiers change the real number substantially.
All three apply across Sonnet 5’s full 1M-token context window with no long-context surcharge — a genuine advantage over Gemini 3.1 Pro, which doubles its rate above 200K tokens.
- Prompt caching: a 5-minute cache write costs 1.25x the base input rate; a 1-hour cache write costs 2x the base input rate. A cache read costs just 10% of the standard input price, so caching pays for itself after a single read on a 5-minute cache. Anthropic cites up to 90% total cost savings for workloads that reuse context heavily.
- Batch processing: asynchronous batch requests get up to 50% off standard pricing, on top of whatever caching discount already applies.
- US-only inference: setting
inference_geoto route requests exclusively through US infrastructure adds a 1.1x multiplier across every token category. Default global routing uses standard pricing with no multiplier.
Stack these correctly and a heavy-caching production workload can land well below the headline rate. Ignore them, and you’re paying full list price for no reason.
Does Claude Sonnet 5 Pricing Actually Beat Opus 4.8 Per Task?
Not always, and this is the part the launch framing glosses over. Anthropic’s headline comparison is per-token: Sonnet 5’s standard rate ($3/$15) is roughly 40% below Opus 4.8’s ($5/$25).
But per-token price and per-task cost are different measurements. Sonnet 5 supports selectable reasoning effort levels from low up to x-high, and higher effort levels burn substantially more tokens to complete the same task.
Third-party benchmark aggregator Artificial Analysis measured Sonnet 5 at roughly $2.29 per task on its Intelligence Index.
At high-effort settings, without introductory pricing in effect, Sonnet 5 can cost more per completed task than Opus 4.8, because the extra reasoning tokens outweigh the lower per-token rate.
The fix isn’t avoiding Sonnet 5. It’s setting an explicit effort-level policy per workload instead of defaulting to the highest setting for everything, and reserving Opus 4.8 for tasks that genuinely need its top-tier accuracy.
Claude Sonnet 5 Pricing vs. GPT-5.6 and Gemini — Live Comparison
All figures below are USD per million tokens.
| Model | Input / 1M | Output / 1M | Notes |
|---|---|---|---|
| Claude Sonnet 5 ENDS AUG 31 | $2.00 | $10.00 | Introductory, through Aug 31 — default Claude model |
| Claude Sonnet 5 FROM SEP 1 | $3.00 | $15.00 | Standard, from Sep 1 — same rate as Sonnet 4.6, tokenizer bills more tokens |
| Claude Opus 4.8 | $5.00 | $25.00 | Highest-accuracy Claude tier |
| Claude Fable 5 NEW | $10.00 | $50.00 | Credit billing, no subscription bundling as of July 8 |
| GPT-5.6 Sol NEW | $5.00 | $30.00 | Top GPT-5.6 preview tier |
| GPT-5.6 Terra NEW | $2.50 | $15.00 | Mid GPT-5.6 preview tier |
| GPT-5.6 Luna NEW | $1.00 | $6.00 | Cheapest GPT-5.6 preview tier |
| GPT-5.5 | $5.00 | $30.00 | Cached input $0.50 |
| Gemini 3.5 Flash | $1.50 | $9.00 | Global endpoint; non-global slightly higher |
| Gemini 3.1 Pro | $2.00 | $12.00 | Up to 200K context; doubles above 200K |
| DeepSeek V4 Flash | $0.14 | $0.28 | Cache miss; cache hit input drops to $0.0028 |
Source: official Anthropic pricing documentation and other providers’ pricing pages, cross-checked as of July 2026. Introductory and preview rates are flagged; confirm live figures before budgeting a production workload.
Calculate Your Own Claude Sonnet 5 Pricing
A rate table tells you the price per token. It doesn’t tell you what your specific workload costs this month, or what it costs after August 31.
That’s a different calculation for every team, depending on model mix, prompt length, caching hit rate, and output volume.
Rather than working through that math by hand, run your own token volume through the calculator below to see July pricing and September pricing side by side.
Frequently Asked Questions
Why did Claude Sonnet 5 get more expensive if the token price looks the same?
When does Claude Sonnet 5’s introductory pricing end?
What changed with Claude Fable 5 pricing?
Is GPT-5.6 or Claude Sonnet 5 cheaper?
Does prompt caching change Claude Sonnet 5 pricing?
Can Claude Sonnet 5 cost more than Opus 4.8 per task?
How This Page Stays Current
The comparison table on this page is generated from a single pricing dataset maintained alongside this article, cross-checked against each provider’s official pricing page.
When a provider changes a rate, this page is updated at the data level rather than rewritten — so the table, the FAQ figures, and the countdown all move together instead of drifting out of sync with each other.
Written and verified by R.K., Creator & Business Economics Analyst
Disclaimer: Pricing figures are sourced from official provider pricing pages and are accurate as of the date shown above.
AI providers change pricing frequently; confirm current rates on the provider’s official pricing page before committing a production budget. Ultimate Info Guide is not affiliated with Anthropic, OpenAI, or Google.
Run Your Own Numbers
This page covers the rates. The calculators do your specific math.
