⚡ Updated Live — July 2026

Claude Sonnet 5 Pricing 2026: Real Rates & Hidden Costs

Current Claude Sonnet 5 pricing, the August 31 rate change, Fable 5’s move to credit billing, and where GPT-5.6 and Gemini 3.5 land in the current price war — with the exact math for your own usage.

Quick Answer

Current Claude Sonnet 5 pricing is $2 per million input tokens and $10 per million output tokens, in effect through August 31, 2026 (roughly seven weeks from now).

After that date, Claude Sonnet 5 pricing rises to standard rates of $3/$15 per million tokens — the same rate Sonnet 4.6 already charged.

The catch: Anthropic confirms Sonnet 5’s new tokenizer produces approximately 30% more tokens for the same text, so a workload with an unchanged rate card can still cost more.

Prompt caching (up to 90% savings) and batch processing (up to 50% savings) apply at full value across the entire 1M-token context window.

Third-party testing from Artificial Analysis found that at high reasoning-effort settings, Sonnet 5 can even cost more per task than Opus 4.8 once introductory pricing ends. Full rate table, cache pricing, and the real budgeting math are below.

Sonnet 5’s tokenizer adds ~30% more tokens — Anthropic’s own figure, with independent testing measuring up to 35% depending on content type
Caching cuts cost up to 90% — cache writes cost 1.25x (5-min) or 2x (1-hour) the base input rate; cache reads cost 10% of standard input price
Sonnet 5 can cost more than Opus 4.8 per task — Artificial Analysis found high-effort Sonnet 5 runs can exceed Opus 4.8’s per-task cost once intro pricing ends
Full 1M context at standard rate — no long-context surcharge, unlike Gemini 3.1 Pro which doubles pricing above 200K tokens

Calculating days remaining… until Claude Sonnet 5’s introductory pricing ends and standard $3/$15 rates apply on August 31, 2026.

How to Use This Page

1

Check what changed

Three separate pricing events, covered below in the order they happened.

2

Scan the live comparison table

Every current model’s rate, side by side, sourced from official pricing pages.

3

Watch the August 31 cliff

Budget at standard pricing now, not just the introductory rate.

4

Run your own numbers

Use the AI Token Calculator for your specific volume and model mix.

What Changed in Claude Sonnet 5 Pricing This Month

Three distinct pricing events landed inside two weeks of each other. Most coverage treats them as separate stories.

They’re not. Together they’re a genuine repricing of how teams should budget AI spend for the second half of 2026.

Anthropic released Claude Sonnet 5 on June 30, 2026 and made it the default model for every Free and Pro user starting July 1. Current Claude Sonnet 5 pricing is an introductory $2 per million input tokens and $10 per million output tokens.

That’s a genuine discount off Sonnet 4.6’s standard rate, but it has an expiration date. On August 31, 2026, the rate steps up to $3/$15 — identical to what Sonnet 4.6 already cost.

A week later, on July 8, 2026, Claude Fable 5 moved to pure usage-credit billing at $10 per million input tokens and $50 per million output tokens.

It’s no longer included in any subscription tier; every Pro, Max, Team, and Business user now draws down purchased credits to use it.

Meanwhile OpenAI’s GPT-5.6 entered preview with three tiers — Sol, Terra, and Luna — deliberately spanning a wide price range. Luna undercuts Sonnet 5’s standard rate; Sol prices above it at every level.

The Hidden Part of Claude Sonnet 5 Pricing: Tokens, Not Just Rates

Here’s the detail most launch-day coverage skipped. Sonnet 5’s standard per-token price is unchanged from Sonnet 4.6, at $3/$15.

But its tokenizer segments the same text into finer-grained units. Anthropic’s own Sonnet model documentation confirms the new tokenizer produces approximately 30% more tokens for identical text, with independent testing measuring as high as 35% on code and structured data.

That’s the real story behind this month’s rate card: it looks flat, but still bills more, because the unit being counted changed underneath it.

A workload that looks cost-neutral this week can land 20–30% above your old baseline once the September rate step-up and the tokenizer inflation are counted together.

The practical takeaway: if you’re migrating a production workload during the introductory window, budget Claude Sonnet 5 pricing against the September 1 standard rate and Anthropic’s confirmed ~30% tokenizer inflation — not the July sticker price. The discount is real, but it’s temporary.

How Caching and Batch Processing Change Claude Sonnet 5 Pricing

The per-token rate is only part of what determines an actual Sonnet 5 bill. Three modifiers change the real number substantially.

All three apply across Sonnet 5’s full 1M-token context window with no long-context surcharge — a genuine advantage over Gemini 3.1 Pro, which doubles its rate above 200K tokens.

  • Prompt caching: a 5-minute cache write costs 1.25x the base input rate; a 1-hour cache write costs 2x the base input rate. A cache read costs just 10% of the standard input price, so caching pays for itself after a single read on a 5-minute cache. Anthropic cites up to 90% total cost savings for workloads that reuse context heavily.
  • Batch processing: asynchronous batch requests get up to 50% off standard pricing, on top of whatever caching discount already applies.
  • US-only inference: setting inference_geo to route requests exclusively through US infrastructure adds a 1.1x multiplier across every token category. Default global routing uses standard pricing with no multiplier.

Stack these correctly and a heavy-caching production workload can land well below the headline rate. Ignore them, and you’re paying full list price for no reason.

Does Claude Sonnet 5 Pricing Actually Beat Opus 4.8 Per Task?

Not always, and this is the part the launch framing glosses over. Anthropic’s headline comparison is per-token: Sonnet 5’s standard rate ($3/$15) is roughly 40% below Opus 4.8’s ($5/$25).

But per-token price and per-task cost are different measurements. Sonnet 5 supports selectable reasoning effort levels from low up to x-high, and higher effort levels burn substantially more tokens to complete the same task.

Third-party benchmark aggregator Artificial Analysis measured Sonnet 5 at roughly $2.29 per task on its Intelligence Index.

At high-effort settings, without introductory pricing in effect, Sonnet 5 can cost more per completed task than Opus 4.8, because the extra reasoning tokens outweigh the lower per-token rate.

The fix isn’t avoiding Sonnet 5. It’s setting an explicit effort-level policy per workload instead of defaulting to the highest setting for everything, and reserving Opus 4.8 for tasks that genuinely need its top-tier accuracy.

Claude Sonnet 5 Pricing vs. GPT-5.6 and Gemini — Live Comparison

All figures below are USD per million tokens.

ModelInput / 1MOutput / 1MNotes
Claude Sonnet 5 ENDS AUG 31$2.00$10.00Introductory, through Aug 31 — default Claude model
Claude Sonnet 5 FROM SEP 1$3.00$15.00Standard, from Sep 1 — same rate as Sonnet 4.6, tokenizer bills more tokens
Claude Opus 4.8$5.00$25.00Highest-accuracy Claude tier
Claude Fable 5 NEW$10.00$50.00Credit billing, no subscription bundling as of July 8
GPT-5.6 Sol NEW$5.00$30.00Top GPT-5.6 preview tier
GPT-5.6 Terra NEW$2.50$15.00Mid GPT-5.6 preview tier
GPT-5.6 Luna NEW$1.00$6.00Cheapest GPT-5.6 preview tier
GPT-5.5$5.00$30.00Cached input $0.50
Gemini 3.5 Flash$1.50$9.00Global endpoint; non-global slightly higher
Gemini 3.1 Pro$2.00$12.00Up to 200K context; doubles above 200K
DeepSeek V4 Flash$0.14$0.28Cache miss; cache hit input drops to $0.0028

Source: official Anthropic pricing documentation and other providers’ pricing pages, cross-checked as of July 2026. Introductory and preview rates are flagged; confirm live figures before budgeting a production workload.

Calculate Your Own Claude Sonnet 5 Pricing

A rate table tells you the price per token. It doesn’t tell you what your specific workload costs this month, or what it costs after August 31.

That’s a different calculation for every team, depending on model mix, prompt length, caching hit rate, and output volume.

Rather than working through that math by hand, run your own token volume through the calculator below to see July pricing and September pricing side by side.

Frequently Asked Questions

Why did Claude Sonnet 5 get more expensive if the token price looks the same?
Sonnet 5’s standard rate ($3 input / $15 output per million tokens) is identical to Sonnet 4.6’s rate. What changed is the tokenizer: it now splits the same text into roughly 1.0 to 1.35 times more tokens. The per-token price held flat while the number of tokens billed per request went up, so the effective cost of unchanged workloads can rise even though the rate card looks the same.
When does Claude Sonnet 5’s introductory pricing end?
Sonnet 5 launched July 1, 2026 at an introductory rate of $2 per million input tokens and $10 per million output tokens. That rate is scheduled to end August 31, 2026, after which standard pricing of $3 per million input tokens and $15 per million output tokens applies.
What changed with Claude Fable 5 pricing?
As of July 8, 2026, Claude Fable 5 moved to pure usage-credit billing at $10 per million input tokens and $50 per million output tokens. It is no longer bundled into any Claude subscription tier; all Fable 5 usage, including on Pro and Max plans, draws from purchased credits.
Is GPT-5.6 or Claude Sonnet 5 cheaper?
It depends on the tier. GPT-5.6 Luna, at roughly $1 input and $6 output per million tokens, undercuts Sonnet 5’s standard pricing. GPT-5.6 Terra, at about $2.50 input and $15 output, sits close to Sonnet 5’s standard rate. GPT-5.6 Sol, at $5 input and $30 output, is priced above Sonnet 5 at every tier. The right comparison depends on which capability tier you actually need.
Does prompt caching change Claude Sonnet 5 pricing?
Yes, substantially. A 5-minute cache write costs 1.25 times the base input rate, and a 1-hour cache write costs 2 times the base input rate. A cache read costs 10% of the standard input price. Anthropic states prompt caching can cut costs by up to 90% and batch processing by up to 50% for eligible workloads, and both apply at full value across Sonnet 5’s entire 1M-token context window.
Can Claude Sonnet 5 cost more than Opus 4.8 per task?
Yes, in some cases. Third-party benchmarking from Artificial Analysis measured Sonnet 5 at roughly $2.29 per task on its Intelligence Index and found that at high reasoning-effort settings, without introductory pricing, Sonnet 5 can cost more per completed task than Opus 4.8 despite its lower per-token rate, because higher effort levels burn substantially more tokens per task.

How This Page Stays Current

The comparison table on this page is generated from a single pricing dataset maintained alongside this article, cross-checked against each provider’s official pricing page.

When a provider changes a rate, this page is updated at the data level rather than rewritten — so the table, the FAQ figures, and the countdown all move together instead of drifting out of sync with each other.

Written and verified by R.K., Creator & Business Economics Analyst

Disclaimer: Pricing figures are sourced from official provider pricing pages and are accurate as of the date shown above.

AI providers change pricing frequently; confirm current rates on the provider’s official pricing page before committing a production budget. Ultimate Info Guide is not affiliated with Anthropic, OpenAI, or Google.

Scroll to Top