Price per 1M logo
2026-09-21 → 2026-09-27 · UTC

AI model price report — week 39, 2026

93 price movements across 5 providers. This report is generated from the same dataset that powers the JSON API and the raw changelog2,045 priced chat models across 33 providers. Dataset refreshed 2026-09-21 00:12 UTC.

Priced chat models
2,045
2,143 endpoints total
Cheapest input
$0.01
Qwen2.5 Coder 7B · Nebius
Median input price
$0.6
half of tracked models are cheaper
Models under $1/1M in
1,296
cheapest tier on the market

Price movements this week

ModelProviderFieldBeforeAfterChange
Gemini ProGoogle Geminiinput per 1m$1.25$2+60.0%
Gemini ProGoogle Geminioutput per 1m$10$12+20.0%
Gemini ProGoogle Geminicached input per 1m$0.125$0.2+60.0%
Gemini Robotics Er 2 PreviewGoogle Geminiinput per 1m$2$1-50.0%
Gemini Robotics Er 2 PreviewGoogle Geminioutput per 1m$10$5-50.0%
Gemini Robotics Er 2 PreviewGoogle Geminicached input per 1m$0.2$0.1-50.0%
Gemini FlashGoogle Geminiinput per 1m$0.3$0.75+150.0%
Gemini FlashGoogle Geminioutput per 1m$2.5$3.75+50.0%
Gemini FlashGoogle Geminicached input per 1m$0.03$0.075+150.0%
Gemini Flash LiteGoogle Geminiinput per 1m$0.1$0.3+200.0%
Gemini Flash LiteGoogle Geminioutput per 1m$0.4$2.5+525.0%
Gemini Flash LiteGoogle Geminicached input per 1m$0.01$0.03+200.0%
GPT 5.6 SolAzure OpenAIinput per 1m$5$4-20.0%
GPT 5.6 SolAzure OpenAIoutput per 1m$30$20-33.3%
GPT 5.6 SolAzure OpenAIcached input per 1m$0.5$0.4-20.0%
GPT 5.1Azure OpenAIinput per 1m$1.38$1.38-0.4%
GPT 5.1Azure OpenAIcached input per 1m$0.14$0.138-1.8%
GPT 5.1 ChatAzure OpenAIinput per 1m$1.38$1.38-0.4%
GPT 5.1 ChatAzure OpenAIcached input per 1m$0.14$0.138-1.8%
GPT 5.1 CodexAzure OpenAIinput per 1m$1.38$1.38-0.4%
GPT 5.1 CodexAzure OpenAIcached input per 1m$0.14$0.138-1.8%
O1 MiniAzure OpenAIinput per 1m$1.21$1.1-9.1%
O1 MiniAzure OpenAIoutput per 1m$4.84$4.4-9.1%
O1 MiniAzure OpenAIcached input per 1m$0.605$0.55-9.1%
GPT 5.1 Codex MiniAzure OpenAIcached input per 1m$0.028$0.028-1.8%
Deepseek V4 ProFireworks AIinput per 1m$1.74$1.2-31.0%
Deepseek V4 ProFireworks AIoutput per 1m$3.48$1.2-65.5%
Deepseek V4 ProFireworks AIcached input per 1m$0.145$0.6+313.8%
Kimi K3OpenRouterinput per 1m$2.1$1.7-19.1%
Kimi K3OpenRouteroutput per 1m$10.53$8.5-19.3%
Kimi K3OpenRoutercached input per 1m$0.235$0.17-27.7%
Deepseek V4 Pro 0813OpenRouterinput per 1m$0.579$1.32+127.8%
Deepseek V4 Pro 0813OpenRouteroutput per 1m$1.74$3.96+127.8%
Deepseek V4 Pro 0813OpenRoutercached input per 1m$0.019$0.044+127.8%
Glm 5.3OpenRouterinput per 1m$1.4$0.91-35.0%
Glm 5.3OpenRouteroutput per 1m$4.4$2.86-35.0%
Glm 5.3OpenRoutercached input per 1m$0.14$0.169+20.7%
Kimi K2.7 CodeOpenRouterinput per 1m$0.71$0.706-0.5%
Kimi K2.7 CodeOpenRouteroutput per 1m$3.5$3.21-8.3%
Kimi K2.7 CodeOpenRoutercached input per 1m$0.15$0.18+20.0%
Glm 5.2OpenRouterinput per 1m$0.6$0.65+8.3%
Glm 5.2OpenRouteroutput per 1m$2$2.04+2.1%
Glm 5.2OpenRoutercached input per 1m$0.15$0.121-19.6%
Nemotron 3 Ultra 550b A55bOpenRouterinput per 1m$0.625$0.6-4.0%
Nemotron 3 Ultra 550b A55bOpenRouteroutput per 1m$3.13$2.4-23.2%
Nemotron 3 Ultra 550b A55bOpenRoutercached input per 1m$0.188$0.12-36.0%
Mistral Large 2512OpenRouterinput per 1m$0.5$0.55+10.0%
Mistral Large 2512OpenRouteroutput per 1m$1.5$1.65+10.0%
Deepseek V4 ProOpenRouterinput per 1m$0.86$0.422-50.9%
Deepseek V4 ProOpenRouteroutput per 1m$1.72$0.845-50.9%
Deepseek V4 ProOpenRoutercached input per 1m$0.072$0.035-50.9%
Minimax M1OpenRouterinput per 1m$0.55$0.4-27.3%
Remm Slerp L2 13bOpenRouterinput per 1m$0.45$0.35-22.2%
Deepseek V4.1 FlashOpenRouterinput per 1m$0.15$0.3+100.0%
Deepseek V4.1 FlashOpenRouteroutput per 1m$0.6$1.2+100.0%
Deepseek V4.1 FlashOpenRoutercached input per 1m$0.003$0.006+100.0%
Qwen3.8 27bOpenRouterinput per 1m$0.42$0.2-52.4%
Qwen3.8 27bOpenRouteroutput per 1m$3$2.55-15.0%
Llama 4 MaverickOpenRouterinput per 1m$0.2$0.188-6.3%
Llama 4 MaverickOpenRouteroutput per 1m$0.696$0.652-6.3%
GPT Oss 120bOpenRouterinput per 1m$0.037$0.15+305.4%
GPT Oss 120bOpenRouteroutput per 1m$0.17$0.6+252.9%
Qwen3.6 35b A3bOpenRouterinput per 1m$0.1$0.15+50.0%
Qwen3.6 35b A3bOpenRouteroutput per 1m$0.9$1+11.1%
Qwen3 Vl 30b A3b InstructOpenRouterinput per 1m$0.15$0.13-13.3%
Qwen3 Vl 30b A3b InstructOpenRouteroutput per 1m$0.6$0.52-13.3%
Qwen3 14bOpenRouterinput per 1m$0.228$0.12-47.3%
Qwen3 14bOpenRouteroutput per 1m$0.91$0.24-73.6%
Mistral Small 3.2 24b InstructOpenRouterinput per 1m$0.075$0.094+25.0%
Mistral Small 3.2 24b InstructOpenRouteroutput per 1m$0.2$0.25+25.0%
Glm 5.3 FlashOpenRouterinput per 1m$0.15$0.09-40.0%
Glm 5.3 FlashOpenRouteroutput per 1m$0.5$0.3-40.0%
Glm 5.3 FlashOpenRoutercached input per 1m$0.03$0.018-40.0%
Gemma 4 26b A4b ItOpenRouterinput per 1m$0.042$0.09+114.3%
Gemma 4 26b A4b ItOpenRouteroutput per 1m$0.22$0.3+36.4%
Qwen3 235b A22b 2507OpenRouterinput per 1m$0.22$0.087-60.2%
Qwen3 235b A22b 2507OpenRouteroutput per 1m$0.88$0.35-60.2%
Mythomax L2 13bOpenRouterinput per 1m$0.06$0.08+33.3%
Mythomax L2 13bOpenRouteroutput per 1m$0.06$0.11+83.3%
Nemotron 3 Super 120b A12bOpenRouterinput per 1m$0.085$0.08-5.9%
Nemotron 3 Super 120b A12bOpenRouteroutput per 1m$0.4$0.45+12.5%
Nemotron 3.5 LightningOpenRouterinput per 1m$0.08$0.07-12.5%
Glm 4.7 FlashOpenRouterinput per 1m$0.06$0.06+0.8%
Nemotron 3 Nano 30b A3bOpenRouterinput per 1m$0.05$0.06+20.0%
Nemotron 3 Nano 30b A3bOpenRouteroutput per 1m$0.2$0.24+20.0%
Qwen3 30b A3b Instruct 2507OpenRouterinput per 1m$0.09$0.048-46.5%
Qwen3 30b A3b Instruct 2507OpenRouteroutput per 1m$0.3$0.193-35.6%
Deepseek V4 Flash 0731OpenRouterinput per 1m$0.065$0.04-38.5%
Deepseek V4 Flash 0731OpenRouteroutput per 1m$0.18$0.16-11.1%
Deepseek V4 FlashOpenRouterinput per 1m$0.085$0.036-58.4%
Deepseek V4 FlashOpenRouteroutput per 1m$0.171$0.071-58.4%
Deepseek V4 FlashOpenRoutercached input per 1m$0.017$0.0071-58.4%
Gemini 2.5 FlashReplicateinput per 1m$2.5$0.3-88.0%
Biggest cuts
  • Gemini 2.5 Pro Preview Tts-99.2%
  • Qwen3 Next 80b A3b Instruct-93.1%
  • Gemini 2.5 Flash-88.0%
  • Glm 4.6-87.5%
  • Mistral Small 3.2 24b Instruct-87.2%
Biggest increases
  • Deepseek Chat V3 0324+1700.0%
  • Gemma 4 26b A4b It+1340.0%
  • Nemotron 3 Super 120b A12b+1340.0%
  • Mistral Large+1150.2%
  • Gemini Omni 1.1 Flash+700.0%

Models added

Models retired

No tracked endpoints disappeared from the upstream sources this week.

Cheapest input tokens

The floor of the market as of this data refresh — what the input side of a bill costs per 1M tokens.

#ModelProviderInput $/1MOutput $/1MCached $/1MContext
1Qwen2.5 Coder 7BNebius$0.01$0.0333K
2Qwen2.5 Coder 3B InstructNscale$0.01$0.03
3Qwen2.5 Coder 7B InstructNscale$0.01$0.03
4Nemotron 3.5 Lightning 30b A3bPerplexity$0.011$0.17$0.0011
5Granite 4.0 H MicroCloudflare Workers AI$0.017$0.112131K
6Granite 4.0 H MicroOpenRouter$0.017$0.112131K
7Mistral Nemo Instruct 2407DeepInfra$0.019$0.03131K
8Mistral NemoOpenRouter$0.019$0.03131K
9Llama 3.2 3B InstructDeepInfra$0.02$0.02131K
10Meta Llama 3.1 8B Instruct TurboDeepInfra$0.02$0.04131K
11Gemma 4 E4B ItDeepInfra$0.02$0.1131K
12Llama Guard 3 8BNebius$0.02$0.06128K
12 models shown — estimate your monthly bill →

Best blended cost

Weighted 3:1 toward input tokens, which is how most chat and agent workloads actually bill.

#ModelProviderInput $/1MOutput $/1MCached $/1MContext
1Gemini Exp 1114Google Geminifreefree1.0M
2Gemini Exp 1206Google Geminifreefree2.1M
3Gemma 3 27b ItGoogle Geminifreefree131K
4Gemma 4 26b A4b ItGoogle Geminifreefree262K
5Gemma 4 31b ItGoogle Geminifreefree262K
6Learnlm 1.5 Pro ExperimentalGoogle Geminifreefree33K
7Labs Leanstral 1 5Mistral AIfreefree262K
8Labs Leanstral 1 5 1Mistral AIfreefree262K
9Anthropic.claude Mythos PreviewAmazon Bedrockfreefree1M
10Gemma 2b It LoraCloudflare Workers AIfreefree8K
11Mistral 7b Instruct V0.2 LoraCloudflare Workers AIfreefree15K
12Llama 2 7b Chat Hf LoraCloudflare Workers AIfreefree8K
12 models shown — estimate your monthly bill →

Most expensive models tracked

The top of the market, for reference — what frontier pricing looks like against the floor above.

#ModelProviderInput $/1MOutput $/1MCached $/1MContext
1O1 ProOpenRouter$150$600200K
2O1 ProOpenAI$150$600200K
3O1 ProOpenAI$150$600200K
4GPT 4.5 PreviewAzure OpenAI$75$150$37.5128K
5GPT 4 32k 0613Azure OpenAI$60$12033K
6GPT 4 32kAzure OpenAI$60$12033K
7GPT 5.4 ProOpenRouter$30$1801.1M
8GPT 5.5 ProOpenRouter$30$1801.1M

What this costs in practice

Three common workloads priced at the cheapest rate available in the index this week.

WorkloadPriced atTotal
support bot · 10,000 chats/day for a month (1.2k in, 400 out)Qwen2.5 Coder 7B · $0.01/$0.03$7.20
1,000,000 document summaries (8k in, 300 out)Qwen2.5 Coder 7B · $0.01/$0.03$89.00
RAG index · 100M embedding tokensPplx Embed V1 0.6b · $0.004/free$0.40

Straight arithmetic from published list prices — no volume discounts, batch tiers or committed-use pricing applied. Run your own numbers in the cost calculator.

Provider price floors

Median and floor input price per provider across the models we track, and how many moved this week.

ProviderModelsFloor inMedian inTop inMoved this week
OpenRouter462$0.017$0.5$150
Fireworks AI273$0.05$0.22$4.5
DeepInfra135$0.019$0.27$16.5
Novita AI132$0.02$0.28$4
OpenAI114$0.05$2$150
Azure OpenAI111$0.05$2$75
Amazon Bedrock109$0.042$0.99$18.8
Together AI100$0.02$0.525$3.5
Mistral AI76$0.06$0.4$4
Databricks67$0.05$1.36$30
Perplexity66$0.011$1.25$10
Nebius58$0.01$0.15$3
Google Gemini49$0.075$0.5$2
xAI44$1$1.25$2
Replicate40$0.03$0.65$15
Cloudflare Workers AI30$0.017$0.351$1.92
Ollama29
Anthropic28$0.25$5$15
Alibaba DashScope27$0.05$0.4$2.5
Moonshot AI24$0.2$1$3
Largest context window
Gemini Exp 1206
2.1M
Free-tier models
76
published at $0 for input tokens
Tracked providers
33
with 2,045 priceable chat models

Frequently asked

What does the 2026-w39 report measure?

It summarises every price difference detected between consecutive refresh runs of the dataset. Each refresh snapshots input, output and cached-input rates per 1M tokens for every tracked endpoint, then diffs that snapshot against the previous one; the movements listed on this page are those diffs, grouped by ISO week.

Why does a week sometimes show no changes?

Because providers do not change published rates every week. A report with zero movements means every diff run in that week came back clean — the figures in the market snapshot sections still reflect the current dataset.

Can I reuse this report?

Yes. The underlying data is published under CC BY 4.0 and served free at /api/models. Attribute Price per 1M with a link when you republish figures.