Home/Models/Mistral Large 3
Verified against Mistral AI Official Pricing
Mistral AIflagshipGenerally Available

Mistral Large 3 Token Counter & Cost Calculator

Mistral flagship reasoning and coding model with aggressive 2026 pricing ($0.50/M in)

Standard Input / 1M
$0.50
Prompt tokens
Cached Input / 1M
$0.10
Prompt caching rate
Output / 1M
$1.50
Generation completion
Context Window
128k
Max out: 16.4k

Interactive Cost & Token Simulator for Mistral Large 3

Live Calculation
Input Tokens2,500
Output Tokens800
Requests / Day5,000
Prompt Caching Ratio (50%)$0.1/M cached rate
Cost / Request$0.00195
Daily Spend (5,000 reqs)$9.75
Monthly Run-Rate (30d)$292.50

Standard Workload Cost Scenarios

ScenarioInput TokensOutput TokensUncached CostWith Prompt Caching
Short Chat Query1,000500$0.00125$0.00085
Document Summarization10,0002,000$0.00800$0.00400
Codebase & Context Analysis100,00020,000$0.0800$0.0400
Batch Corpus Processing1,000,000100,000$0.6500$0.2500
Technical Architecture & Pricing VerificationVerified: September 8, 2026
Model ArchitectureMistral Large 3
Provider OrganizationMistral AI
Tokenizer EncodingMistral Tekken Tokenizer (131k vocabulary)
API ConnectivityMistral La Plateforme API (/v1/chat/completions) & Azure AI Studio.
Batch API 50% DiscountSupported (50% off input & output)
Tier 1 Rate Quota300 RPM · 1M TPM
Blended 3:1 Cost / 1M Tokens$0.75
Mistral flagship model with native European data sovereignty and multilingual strength.Mistral AI Official Pricing

Verified Pricing History

August 2026: $2/M input · $6/M output ($0.5/M cached)
Mistral Large 3 release.

When to Choose Mistral Large 3

Ideal for production workloads demanding flagship capabilities, deep context depth (128k tokens), and reliability from Mistral AI. Excellent when predictable tokenomics and prompt caching support are paramount.

When Another Model May Be Better

If your use-case requires sub-second streaming latency or ultra-high frequency classification at micro-cent pricing, consider lighter budget options such as Gemini Flash-Lite or Claude Haiku. For deep formal logic, consider dedicated reasoning models like o3.

Frequently Asked Questions About Mistral Large 3

How much does 1 million tokens cost with Mistral Large 3?

For Mistral Large 3, 1 million input tokens costs $0.50, while 1 million output tokens costs $1.50. If using prompt caching, repetitive input prefixes are discounted to $0.10 per million.

What is the context window for Mistral Large 3?

Mistral Large 3 features a maximum context window of 128,000 tokens (~96,000 words), with a maximum output limit of 16,384 tokens per completion.

Which tokenizer does Mistral Large 3 use?

Mistral Large 3 utilizes the Mistral Tekken Tokenizer (131k vocabulary). Token counting on TokenMath runs client-side to ensure maximum privacy.

Does Mistral Large 3 support prompt caching discounts?

Yes. Mistral Large 3 supports prompt caching with a cached input rate of $0.10/1M (saving up to 80% on repeated input context).