Gemini 2.5 Flash-Lite alternatives

5 models of comparable capability, ranked by blended token price. Gemini 2.5 Flash-Lite stays in the table, highlighted, as the baseline.

List prices checked

Ranked by blended token price

Cheapest first, at three input tokens per output token.

#ModelProviderInputOutputBlendedvs Gemini 2.5 Flash-LiteContext
1GPT-5 nanoOpenAI$0.050$0.400$0.1381.3× cheaper400K
2Mistral SmallMistral AI$0.100$0.300$0.1501.2× cheaper131K
3Gemini 2.5 Flash-Liteyou are hereGoogle$0.100$0.400$0.1751M
4GPT-4o miniOpenAI$0.150$0.600$0.2621.4× dearer128K
5Cohere Command RCohere$0.150$0.600$0.2621.4× dearer128K
6GPT-4.1 miniOpenAI$0.400$1.60$0.7003.3× dearer1M

Source: Google pricing (May 2026). Published list prices only — committed-use discounts, regional variation, and per-request charges are not included. Not yet independently re-checked against the vendor’s page.

Questions

What is the best alternative to Gemini 2.5 Flash-Lite in 2026?
On blended token price, GPT-5 nano is the cheapest comparable option — about 1.3× cheaper than Gemini 2.5 Flash-Lite. Price is only half the decision: a model that costs 70% less but needs a second attempt on hard prompts is not a saving.
Which alternative is cheapest for output-heavy work?
Rank on the output column rather than the blended figure. Reasoning and agentic loops bill their thinking as output, so a model with cheap input and expensive output can end up the most costly option on the page despite looking competitive here.
Do these alternatives have the same context window?
No — the context column varies by more than 8× across this table. If you are feeding whole documents or codebases, filter on window size first and only then compare price, because a cheaper model that cannot hold your input is not an alternative at all.

Compare Gemini 2.5 Flash-Lite directly