Home  /  Models  /  Mistral Large 2
⚡ Quick Pricing Summary

Mistral Large 2 by Mistral AI costs $2.00 per 1 Million input tokens and $6.00 per 1 Million output tokens. It features a 262k context window and has a blended 3:1 production rate of $3.000/1M tokens.

Mistral AI Flagship (EU Sovereign)

Mistral Large 2

European flagship 123B model with native multilingual fluency (FR, DE, ES, IT, AR, HI, JA, ZH).

Calculate Spend in App →
Input Token Price
$2.00
Per 1,000,000 tokens
Output Token Price
$6.00
Per 1,000,000 tokens
Prompt Caching Rate
Not Supported
Save up to 80% on cached inputs
Max Context Window
262k
Max output: 262k tokens

Verified Specifications & Benchmark Data

Developer / Provider Mistral AI
Model Family Mistral
Blended 3:1 Rate (Production Benchmark) $3.000 / 1M tokens
Batch API Discount (24hr SLA) None
Multimodal Vision ❌ Text Only
Function Calling / Structured Outputs ✅ Native Tool Calling
MMLU Benchmark Score 84.0%
Latency & Throughput Tier Fast (~45 t/s)
Optimal Architecture & Use Cases Multilingual EU business workflows, code generation, reasoning

Frequently Asked Questions about Mistral Large 2

How much does Mistral Large 2 cost per 1M tokens?

Mistral Large 2 pricing is set at $2.00 per 1 million input tokens and $6.00 per 1 million output tokens. For high-volume batch workloads, 24-hour batch queues provide a None.

How much money does Mistral Large 2 prompt caching save?

Prompt caching is currently not natively offered for Mistral Large 2 on direct serverless endpoints.

What is the context window limit of Mistral Large 2?

Mistral Large 2 has a maximum context window of 262k (262,144 tokens), supporting up to 262,144 completion tokens per response.

Direct Matchups with Mistral Large 2

See how Mistral Large 2 compares against other leading frontier and open-weights models in cost and latency.