DeepSeek-V4-Flash alternatives

8 models of comparable capability, ranked by blended token price. DeepSeek-V4-Flash stays in the table, highlighted, as the baseline.

List prices checked

Ranked by blended token price

Cheapest first, at three input tokens per output token.

#ModelProviderInputOutputBlendedvs DeepSeek-V4-FlashContext
1DeepSeek-V4-Flashyou are hereDeepSeek$0.140$0.280$0.1751M
2Grok 3 minixAI$0.300$0.500$0.3502× dearer131K
3CodestralMistral AI$0.300$0.900$0.4502.5× dearer262K
4GPT-5 miniOpenAI$0.250$2.00$0.6883.3× dearer400K
5Gemini 2.5 FlashGoogle$0.300$2.50$0.8505× dearer1M
6OpenAI o4-miniOpenAI$1.10$4.40$1.9310× dearer200K
7Claude Haiku 4.5Anthropic$1.00$5.00$2.0010× dearer200K
8GPT-4.1OpenAI$2.00$8.00$3.5010× dearer1M
9GPT-4oOpenAI$2.50$10.00$4.38Infinity× dearer128K

Source: DeepSeek pricing (August 2026). Published list prices only — committed-use discounts, regional variation, and per-request charges are not included. Not yet independently re-checked against the vendor’s page.

Questions

What is the best alternative to DeepSeek-V4-Flash in 2026?
Nothing in the same capability tier is cheaper than DeepSeek-V4-Flash on a blended basis, which makes it the value pick of its tier. Moving down a tier would cut cost further at some cost in capability.
Which alternative is cheapest for output-heavy work?
Rank on the output column rather than the blended figure. Reasoning and agentic loops bill their thinking as output, so a model with cheap input and expensive output can end up the most costly option on the page despite looking competitive here.
Do these alternatives have the same context window?
No — the context column varies by more than 8× across this table. If you are feeding whole documents or codebases, filter on window size first and only then compare price, because a cheaper model that cannot hold your input is not an alternative at all.

Compare DeepSeek-V4-Flash directly