Price per 1M logo

GPT 5.4 vs Deepseek V4 Flash pricing

Side-by-side token pricing for GPT 5.4 (OpenAI) and Deepseek V4 Flash (DeepSeek), plus the total cost at three workload sizes so the crossover point is obvious. Prices refreshed 2026-09-13 13:40 UTC.

MetricGPT 5.4Deepseek V4 Flash
Input $/1M tokens$2.5$0.3
Output $/1M tokens$15$1.2
Cached input $/1M$0.25$0.006
Context window1.1M1M
Max output128K393K
Visionyesyes
Tool callingyesyes
Cache pricingyesyes

Total cost by workload

WorkloadGPT 5.4Deepseek V4 FlashCheaper
1M tokens in · 300K out$7$0.66Deepseek V4 Flash (91% less)
10M tokens in · 3M out$70$6.6Deepseek V4 Flash (91% less)
100M tokens in · 30M out$700$66Deepseek V4 Flash (91% less)

On blended cost (weighted 3:1 toward input), Deepseek V4 Flash comes out cheaper by roughly 88% on input pricing. Pick Deepseek V4 Flash when the workload is price-sensitive and its context window covers your prompts; GPT 5.4 is the one to benchmark if you need vision, reasoning depth, or a different quality profile.

Full breakdown
GPT 5.4
Full breakdown
Deepseek V4 Flash

Frequently asked

Is GPT 5.4 cheaper than Deepseek V4 Flash?

On input pricing, Deepseek V4 Flash is cheaper ($0.3 vs $2.5 per 1M tokens). On blended cost the winner is Deepseek V4 Flash, because output tokens are weighted differently for most workloads.

Which has the larger context window?

GPT 5.4 offers the larger context (1.1M tokens).

Do these prices include prompt caching?

Cached input is priced separately where a provider publishes a cache rate; the table above lists it per model so the comparison stays like-for-like.