Price per 1M logo

GPT 5.4 vs Gemini 2.5 Flash Lite Preview 09 2025 pricing

Side-by-side token pricing for GPT 5.4 (OpenAI) and Gemini 2.5 Flash Lite Preview 09 2025 (Google Gemini), plus the total cost at three workload sizes so the crossover point is obvious. Prices refreshed 2026-09-18 00:09 UTC.

MetricGPT 5.4Gemini 2.5 Flash Lite Preview 09 2025
Input $/1M tokens$2.5$0.1
Output $/1M tokens$15$0.4
Cached input $/1M$0.25$0.01
Context window1.1M1.0M
Max output128K66K
Visionyesyes
Tool callingyesyes
Cache pricingyesyes

Total cost by workload

WorkloadGPT 5.4Gemini 2.5 Flash Lite Preview 09 2025Cheaper
1M tokens in · 300K out$7$0.22Gemini 2.5 Flash Lite Preview 09 2025 (97% less)
10M tokens in · 3M out$70$2.2Gemini 2.5 Flash Lite Preview 09 2025 (97% less)
100M tokens in · 30M out$700$22Gemini 2.5 Flash Lite Preview 09 2025 (97% less)

On blended cost (weighted 3:1 toward input), Gemini 2.5 Flash Lite Preview 09 2025 comes out cheaper by roughly 96% on input pricing. Pick Gemini 2.5 Flash Lite Preview 09 2025 when the workload is price-sensitive and its context window covers your prompts; GPT 5.4 is the one to benchmark if you need vision, reasoning depth, or a different quality profile.

Full breakdown
GPT 5.4
Full breakdown
Gemini 2.5 Flash Lite Preview 09 2025

Frequently asked

Is GPT 5.4 cheaper than Gemini 2.5 Flash Lite Preview 09 2025?

On input pricing, Gemini 2.5 Flash Lite Preview 09 2025 is cheaper ($0.1 vs $2.5 per 1M tokens). On blended cost the winner is Gemini 2.5 Flash Lite Preview 09 2025, because output tokens are weighted differently for most workloads.

Which has the larger context window?

GPT 5.4 offers the larger context (1.1M tokens).

Do these prices include prompt caching?

Cached input is priced separately where a provider publishes a cache rate; the table above lists it per model so the comparison stays like-for-like.