Price per 1M logo

GPT 5.4 vs Grok 4 1 Fast Non Reasoning pricing

Side-by-side token pricing for GPT 5.4 (OpenAI) and Grok 4 1 Fast Non Reasoning (xAI), plus the total cost at three workload sizes so the crossover point is obvious. Prices refreshed 2026-09-13 13:40 UTC.

MetricGPT 5.4Grok 4 1 Fast Non Reasoning
Input $/1M tokens$2.5$1.25
Output $/1M tokens$15$2.5
Cached input $/1M$0.25$0.2
Context window1.1M2M
Max output128K2M
Visionyesyes
Tool callingyesyes
Cache pricingyesyes

Total cost by workload

WorkloadGPT 5.4Grok 4 1 Fast Non ReasoningCheaper
1M tokens in · 300K out$7$2Grok 4 1 Fast Non Reasoning (71% less)
10M tokens in · 3M out$70$20Grok 4 1 Fast Non Reasoning (71% less)
100M tokens in · 30M out$700$200Grok 4 1 Fast Non Reasoning (71% less)

On blended cost (weighted 3:1 toward input), Grok 4 1 Fast Non Reasoning comes out cheaper by roughly 50% on input pricing. Pick Grok 4 1 Fast Non Reasoning when the workload is price-sensitive and its context window covers your prompts; GPT 5.4 is the one to benchmark if you need vision, reasoning depth, or a different quality profile.

Full breakdown
GPT 5.4
Full breakdown
Grok 4 1 Fast Non Reasoning

Frequently asked

Is GPT 5.4 cheaper than Grok 4 1 Fast Non Reasoning?

On input pricing, Grok 4 1 Fast Non Reasoning is cheaper ($1.25 vs $2.5 per 1M tokens). On blended cost the winner is Grok 4 1 Fast Non Reasoning, because output tokens are weighted differently for most workloads.

Which has the larger context window?

Grok 4 1 Fast Non Reasoning offers the larger context (2M tokens).

Do these prices include prompt caching?

Cached input is priced separately where a provider publishes a cache rate; the table above lists it per model so the comparison stays like-for-like.