Private beta DPDP-ready sovereign routing is live, request access
Mistral

mistral-large pricing

List price for the mistral-large API from Mistral — input cost, output cost, and context window, plus what it actually costs to run across a few common workloads. All figures in USD per 1M tokens.

Last updated: July 2026
Input
$2.00 / 1M tokens
Output
$6.00 / 1M tokens
1M in + 1M out
$8.00 blended
Context window
128K tokens

What mistral-large costs for typical workloads

Estimated cost per request and per 1,000 requests at mistral-large's list price, for a few representative input/output token mixes. Your real usage will vary with prompt and response length.

Workload Input tokens Output tokens Cost / request Cost / 1,000
Short chat turn 1,000 300 $0.00380 $3.80
RAG answer 8,000 500 $0.019 $19.00
Agent step (tool-heavy) 20,000 2,000 $0.052 $52.00
Long-document summary 100,000 1,000 $0.206 $206

Provider-published list price, USD per 1M tokens, current as of July 2026. Batch, cached-input, and volume discounts are not applied. Verify with Mistral before relying on these figures for billing.

Route to the cheapest model automatically.

Stop hand-tuning which model gets which request. Routeplane's difficulty router scores each prompt and picks the optimal model per request — downgrading easy calls to a cheaper model and reserving the expensive ones for the hard prompts, with automatic fallback if a provider is down. One OpenAI-compatible endpoint, every provider behind it.