Mistral Large pricing

The French lab's flagship, priced below every US frontier model and hosted in the EU, which is often the deciding factor rather than the benchmark scores.

List prices checked

By Mistral AI · Frontier tier

Token pricing

Per million tokens, from the official price list.

Token typePrice per 1MNotes
Input$2.00Everything you send, including the system prompt
Output$6.00Everything the model generates, reasoning tokens included
Blended (3:1)$3.00One number for ranking, at three input tokens per output token

Source: Mistral AI pricing (May 2026). Published list prices only — committed-use discounts, regional variation, and per-request charges are not included. Not yet independently re-checked against the vendor’s page.

Capabilities

PropertyValue
Context window131K tokens
Inputs acceptedtext
Cost to fill the context once$0.26

What it costs per month

Token prices applied to three workload shapes. Input-heavy and output-heavy workloads rank models very differently.

In-app chat assistant

A support or help assistant handling a few thousand conversations a month, with short prompts and short answers.

$70/mo

  • 20M input tokens$40
  • 5M output tokens$30

RAG over documents

Retrieval-augmented answering where each request stuffs several retrieved passages into the prompt, so input dominates.

$690/mo

  • 300M input tokens$600
  • 15M output tokens$90

Coding agent

An agentic loop that reads files and writes patches, generating heavy output and re-sending large context on every turn.

$1,400/mo

  • 400M input tokens$800
  • 100M output tokens$600

Mistral Large vs other frontier models

Same capability tier, ranked by blended token price.

ModelInputOutputContextCompare
OpenAI o3$2.00$8.00200KHead to head
Cohere Command A$2.50$10.00256KHead to head
GPT-5$1.25$10.00400KHead to head
Gemini 2.5 Pro$1.25$10.001MHead to head
Claude Sonnet 5$3.00$15.00200KHead to head
Grok 4$3.00$15.00256KHead to head

Questions about pricing

How much does the Mistral Large API cost in 2026?
$2.00 per million input tokens and $6.00 per million output tokens, from Mistral AI's published pricing. The table above converts that into a monthly figure for three common workloads.
What is the Mistral Large context window?
131K tokens. Filling the window costs whatever that many input tokens cost — at $2.00 per million, a single full-context request runs about $0.26.
Is Mistral Large cheaper than the alternatives?
On a blended 3:1 input-to-output basis it works out at $3.00 per million tokens. The comparison table below ranks it against the other frontier models in the dataset, and the head-to-head pages model both against identical workloads.

Keep comparing