Comparison · prices as of 2026-07-04 · re-verified 2026-08-18

GPT-5 mini vs Gemini 2.5 Flash-Lite

On price alone, Gemini 2.5 Flash-Lite (Google) runs about 73% cheaper than GPT-5 mini for a typical chat workload (2K input + 500 output tokens, 1,000 requests/day). Whether that's the right trade depends on the capability gap — the breakdown below shows where each one wins.

Price per million tokens

Spec GPT-5 mini Gemini 2.5 Flash-Lite
Input / MTok $0.25 $0.10
Output / MTok $2 $0.40
Context window 400K 1M
Max output 128K 66K
Class workhorse fast / lightweight

Estimated monthly cost by workload

Same models, three realistic usage levels. The gap between them scales with volume, so the right pick can change as you grow.

Workload GPT-5 mini Gemini 2.5 Flash-Lite Cheaper
Light — chatbot 0.5K in / 0.2K out · 200 req/day $3.15 $0.78 Gemini 2.5 Flash-Lite
Typical — assistant 2K in / 0.5K out · 1K req/day $45.00 $12.00 Gemini 2.5 Flash-Lite
Heavy — RAG / agent 8K in / 2K out · 5K req/day $900 $240 Gemini 2.5 Flash-Lite

Which should you pick?

Frequently asked questions

Is GPT-5 mini or Gemini 2.5 Flash-Lite cheaper?

Gemini 2.5 Flash-Lite is cheaper — roughly 73% less per month on a typical assistant workload (2K input + 500 output tokens, 1,000 requests/day).

What do GPT-5 mini and Gemini 2.5 Flash-Lite cost per million tokens?

GPT-5 mini is $0.25 per million input tokens and $2 per million output. Gemini 2.5 Flash-Lite is $0.10 input and $0.40 output. Prices are first-party API list rates as of 2026-07-04.

Which has the larger context window, GPT-5 mini or Gemini 2.5 Flash-Lite?

Gemini 2.5 Flash-Lite does, at 1M tokens versus 400K for GPT-5 mini.

When should I choose GPT-5 mini over Gemini 2.5 Flash-Lite?

Choose GPT-5 mini when you need the extra capability of its workhorse class; drop to Gemini 2.5 Flash-Lite when speed and cost matter more than peak quality.

Related comparisons

Sources: OpenAI pricing · Google pricing. Run your own numbers in the cost calculator.