Comparison · prices as of 2026-08-30 · re-verified 2026-10-06

Gemini 2.5 Flash vs Gemini 2.5 Flash-Lite

On price alone, Gemini 2.5 Flash-Lite (Google) runs about 78% cheaper than Gemini 2.5 Flash for a typical chat workload (2K input + 500 output tokens, 1,000 requests/day). Whether that's the right trade depends on the capability gap — the breakdown below shows where each one wins.

Price per million tokens

Spec Gemini 2.5 Flash Gemini 2.5 Flash-Lite
Input / MTok $0.30 $0.10
Output / MTok $2.50 $0.40
Context window 1M 1M
Max output 66K 66K
Class workhorse fast / lightweight

Estimated monthly cost by workload

Same models, three realistic usage levels. The gap between them scales with volume, so the right pick can change as you grow.

Workload Gemini 2.5 Flash Gemini 2.5 Flash-Lite Cheaper
Light — chatbot 0.5K in / 0.2K out · 200 req/day $3.90 $0.78 Gemini 2.5 Flash-Lite
Typical — assistant 2K in / 0.5K out · 1K req/day $55.50 $12.00 Gemini 2.5 Flash-Lite
Heavy — RAG / agent 8K in / 2K out · 5K req/day $1,110 $240 Gemini 2.5 Flash-Lite

Which should you pick?

Frequently asked questions

Is Gemini 2.5 Flash or Gemini 2.5 Flash-Lite cheaper?

Gemini 2.5 Flash-Lite is cheaper — roughly 78% less per month on a typical assistant workload (2K input + 500 output tokens, 1,000 requests/day).

What do Gemini 2.5 Flash and Gemini 2.5 Flash-Lite cost per million tokens?

Gemini 2.5 Flash is $0.30 per million input tokens and $2.50 per million output. Gemini 2.5 Flash-Lite is $0.10 input and $0.40 output. Prices are first-party API list rates as of 2026-08-30.

Which has the larger context window, Gemini 2.5 Flash or Gemini 2.5 Flash-Lite?

Both handle up to 1M tokens of context.

When should I choose Gemini 2.5 Flash over Gemini 2.5 Flash-Lite?

Choose Gemini 2.5 Flash when you need the extra capability of its workhorse class; drop to Gemini 2.5 Flash-Lite when speed and cost matter more than peak quality.

Related comparisons

Sources: Google pricing · Google pricing. Run your own numbers in the cost calculator.