Cheapest EU and open-weight model APIs

Avoiding US hyperscaler lock-in, for data residency or licensing reasons.

List prices checked

How this is scored

Models from providers outside the big three US labs, ranked on blended token price. Some are open-weight and can be self-hosted later, which changes the negotiating position even if you start on the hosted API.

All 7, ranked for this workload

#ModelProviderCost for this workloadNotes
1Mistral SmallMistral AI$0.15/1M blendedAn open-weight model available both as a hosted API and as a download, which makes it the cheap tier you can also run yourself if the economics change.
2DeepSeek-V4-FlashDeepSeek$0.18/1M blendedA million-token context window at a fraction of any Western model's price, with cache-hit input billed at a fiftieth of the base rate. DeepSeek has published notice of a significant price rise, so treat these rates as a floor rather than a plan.
3Cohere Command RCohere$0.26/1M blendedThe cheap end of Cohere's range, aimed at retrieval pipelines where the model's job is to ground an answer in supplied documents rather than to reason from scratch.
4CodestralMistral AI$0.45/1M blendedA code-specialised model tuned for fill-in-the-middle completion, priced for the high request volume that inline editor autocomplete generates.
5DeepSeek-V4-ProDeepSeek$0.54/1M blendedFrontier-class reasoning at roughly a fifth of Western frontier pricing, with a million-token window. The trade-offs are Chinese data residency, a 500-request concurrency cap, and a price rise the vendor has already announced.
6Mistral LargeMistral AI$3.00/1M blendedThe French lab's flagship, priced below every US frontier model and hosted in the EU, which is often the deciding factor rather than the benchmark scores.
7Cohere Command ACohere$4.38/1M blendedCohere's enterprise flagship, built around retrieval-augmented generation and multilingual work, and sold heavily as a private deployment inside customer infrastructure.

Published list prices only. Rows that do not publish a rate for an axis this workload depends on are excluded rather than shown as cheap.

Questions

Cheapest EU and open-weight model APIs in 2026?
Mistral Small from Mistral AI — $0.15/1M blended. DeepSeek-V4-Flash is second at $0.18/1M blended. This ranking is specific to the workload described above; a different usage shape reorders it.
How is this ranking weighted?
Models from providers outside the big three US labs, ranked on blended token price. Some are open-weight and can be self-hosted later, which changes the negotiating position even if you start on the hosted API.
Why does this differ from the general cheapest list?
Because a single price axis never describes a real workload. Ranking by headline rate answers "who is cheapest per unit"; this page answers "who is cheapest for this job", and the two orders are often very different — which is the whole reason it exists as a separate table.
What is not accounted for?
Committed-use discounts, regional price variation, per-request charges, support plans, and minimum retention. These are published list prices applied to one stated workload — a shortlist to price properly with the vendor, not a quotation.

Other workloads