Home/Models/Claude Sonnet 5
Verified against Anthropic Official Pricing
AnthropicflagshipGenerally Available

Claude Sonnet 5 Token Counter & Cost Calculator

High-performance frontier model for production coding and complex agents

Standard Input / 1M
$3.00
Prompt tokens
Cached Input / 1M
$0.20
Prompt caching rate
Output / 1M
$15.00
Generation completion
Context Window
200k
Max out: 16.4k

Interactive Cost & Token Simulator for Claude Sonnet 5

Live Calculation
Input Tokens2,500
Output Tokens800
Requests / Day5,000
Prompt Caching Ratio (50%)$0.2/M cached rate
Cost / Request$0.0160
Daily Spend (5,000 reqs)$80.00
Monthly Run-Rate (30d)$2,400.00

Standard Workload Cost Scenarios

ScenarioInput TokensOutput TokensUncached CostWith Prompt Caching
Short Chat Query1,000500$0.0105$0.00770
Document Summarization10,0002,000$0.0600$0.0320
Codebase & Context Analysis100,00020,000$0.6000$0.3200
Batch Corpus Processing1,000,000100,000$4.50$1.70
Technical Architecture & Pricing VerificationVerified: September 8, 2026
Model ArchitectureClaude Sonnet 5
Provider OrganizationAnthropic
Tokenizer EncodingAnthropic Claude BPE (~65k vocabulary)
API ConnectivityAnthropic Messages API (/v1/messages). Native prompt caching and 16k output tokens.
Batch API 50% DiscountSupported (50% off input & output)
Cache Write Rate / 1M$2.50
Tier 1 Rate Quota1,000 RPM · 200k TPM
Blended 3:1 Cost / 1M Tokens$6.00
Most cost-efficient frontier coding model. Prompt cache write rate is $3.75/1M, reads are $0.30/1M.Anthropic Official Pricing

Verified Pricing History

August 2026: $3/M input · $15/M output ($0.2/M cached)
Claude Sonnet 5 release, matching Claude 3.5 Sonnet pricing with enhanced performance.
June 2024: $3/M input · $15/M output ($0.3/M cached)
Claude 3.5 Sonnet baseline pricing introduced.

When to Choose Claude Sonnet 5

Ideal for production workloads demanding flagship capabilities, deep context depth (200k tokens), and reliability from Anthropic. Excellent when predictable tokenomics and prompt caching support are paramount.

When Another Model May Be Better

If your use-case requires sub-second streaming latency or ultra-high frequency classification at micro-cent pricing, consider lighter budget options such as Gemini Flash-Lite or Claude Haiku. For deep formal logic, consider dedicated reasoning models like o3.

Frequently Asked Questions About Claude Sonnet 5

How much does 1 million tokens cost with Claude Sonnet 5?

For Claude Sonnet 5, 1 million input tokens costs $3.00, while 1 million output tokens costs $15.00. If using prompt caching, repetitive input prefixes are discounted to $0.20 per million.

What is the context window for Claude Sonnet 5?

Claude Sonnet 5 features a maximum context window of 200,000 tokens (~150,000 words), with a maximum output limit of 16,384 tokens per completion.

Which tokenizer does Claude Sonnet 5 use?

Claude Sonnet 5 utilizes the Anthropic Claude BPE (~65k vocabulary). Token counting on TokenMath runs client-side to ensure maximum privacy.

Does Claude Sonnet 5 support prompt caching discounts?

Yes. Claude Sonnet 5 supports prompt caching with a cached input rate of $0.20/1M (saving up to 93% on repeated input context).