Price per 1M logo
2026-09-07 → 2026-09-13 · UTC

AI model price report — week 37, 2026

No price movements detected. This report is generated from the same dataset that powers the JSON API and the raw changelog1,809 priced chat models across 33 providers. Dataset refreshed 2026-09-13 13:40 UTC.

Priced chat models
1,809
1,907 endpoints total
Cheapest input
$0.01
Qwen2.5 Coder 7B · Nebius
Median input price
$0.56
half of tracked models are cheaper
Models under $1/1M in
1,143
cheapest tier on the market

Price movements this week

No price change was detected in any refresh run during this week. The dataset is diffed against the previous snapshot on every run, so a quiet week means the providers we track held their published rates. Systematic tracking of price history began on 2026-09-13, so earlier weeks carry no comparison baseline.

1 refresh run this week · 0 total diffs recorded.

Models added

No new endpoints appeared in this week's refresh runs.

Models retired

No tracked endpoints disappeared from the upstream sources this week.

Cheapest input tokens

The floor of the market as of this data refresh — what the input side of a bill costs per 1M tokens.

#ModelProviderInput $/1MOutput $/1MCached $/1MContext
1Qwen2.5 Coder 7BNebius$0.01$0.0333K
2Qwen2.5 Coder 3B InstructNscale$0.01$0.03
3Qwen2.5 Coder 7B InstructNscale$0.01$0.03
4Nemotron 3.5 Lightning 30b A3bPerplexity$0.011$0.17$0.0011
5Granite 4.0 H MicroCloudflare Workers AI$0.017$0.112131K
6Mistral Nemo Instruct 2407DeepInfra$0.019$0.03131K
7Mistral NemoOpenRouter$0.019$0.03131K
8Llama 3.2 3B InstructDeepInfra$0.02$0.02131K
9Meta Llama 3.1 8B Instruct TurboDeepInfra$0.02$0.04131K
10Gemma 4 E4B ItDeepInfra$0.02$0.1131K
11Llama Guard 3 8BNebius$0.02$0.06128K
12Meta Llama 3.1 8B InstructNebius$0.02$0.06128K
12 models shown — estimate your monthly bill →

Best blended cost

Weighted 3:1 toward input tokens, which is how most chat and agent workloads actually bill.

#ModelProviderInput $/1MOutput $/1MCached $/1MContext
1Gemini Exp 1114Google Geminifreefree1.0M
2Gemini Exp 1206Google Geminifreefree2.1M
3Gemma 3 27b ItGoogle Geminifreefree131K
4Gemma 4 26b A4b ItGoogle Geminifreefree262K
5Gemma 4 31b ItGoogle Geminifreefree262K
6Learnlm 1.5 Pro ExperimentalGoogle Geminifreefree33K
7Labs Leanstral 1 5Mistral AIfreefree262K
8Labs Leanstral 1 5 1Mistral AIfreefree262K
9Anthropic.claude Mythos PreviewAmazon Bedrockfreefree1M
10Gemma 2b It LoraCloudflare Workers AIfreefree8K
11Mistral 7b Instruct V0.2 LoraCloudflare Workers AIfreefree15K
12Llama 2 7b Chat Hf LoraCloudflare Workers AIfreefree8K
12 models shown — estimate your monthly bill →

Most expensive models tracked

The top of the market, for reference — what frontier pricing looks like against the floor above.

#ModelProviderInput $/1MOutput $/1MCached $/1MContext
1O1 ProOpenRouter$150$600200K
2O1 ProOpenAI$150$600200K
3O1 ProOpenAI$150$600200K
4GPT 4.5 PreviewAzure OpenAI$75$150$37.5128K
5GPT 4 32k 0613Azure OpenAI$60$12033K
6GPT 4 32kAzure OpenAI$60$12033K
7GPT 5.4 ProOpenRouter$30$1801.1M
8GPT 5.5 ProOpenRouter$30$1801.1M

What this costs in practice

Three common workloads priced at the cheapest rate available in the index this week.

WorkloadPriced atTotal
support bot · 10,000 chats/day for a month (1.2k in, 400 out)Qwen2.5 Coder 7B · $0.01/$0.03$7.20
1,000,000 document summaries (8k in, 300 out)Qwen2.5 Coder 7B · $0.01/$0.03$89.00
RAG index · 100M embedding tokensPplx Embed V1 0.6b · $0.004/free$0.40

Straight arithmetic from published list prices — no volume discounts, batch tiers or committed-use pricing applied. Run your own numbers in the cost calculator.

Provider price floors

Median and floor input price per provider across the models we track, and how many moved this week.

ProviderModelsFloor inMedian inTop inMoved this week
Fireworks AI272$0.05$0.22$4.5
OpenRouter268$0.019$0.45$150
DeepInfra135$0.019$0.27$16.5
Novita AI132$0.02$0.28$4
OpenAI112$0.05$2$150
Amazon Bedrock109$0.042$0.99$18.8
Azure OpenAI104$0.05$1.88$75
Together AI78$0.05$0.5$3.5
Mistral AI73$0.06$0.4$4
Databricks67$0.05$1.36$30
Perplexity66$0.011$1.25$10
Nebius56$0.01$0.15$3
Google Gemini49$0.075$0.35$2
xAI44$1$1.25$2
Replicate40$0.03$0.65$15
Cloudflare Workers AI30$0.017$0.351$1.92
Ollama29
Anthropic28$0.25$5$15
Alibaba DashScope25$0.05$0.4$2.5
Moonshot AI24$0.2$1$3
Largest context window
Gemini Exp 1206
2.1M
Free-tier models
63
published at $0 for input tokens
Tracked providers
33
with 1,809 priceable chat models

Frequently asked

What does the 2026-w37 report measure?

It summarises every price difference detected between consecutive refresh runs of the dataset. Each refresh snapshots input, output and cached-input rates per 1M tokens for every tracked endpoint, then diffs that snapshot against the previous one; the movements listed on this page are those diffs, grouped by ISO week.

Why does a week sometimes show no changes?

Because providers do not change published rates every week. A report with zero movements means every diff run in that week came back clean — the figures in the market snapshot sections still reflect the current dataset.

Can I reuse this report?

Yes. The underlying data is published under CC BY 4.0 and served free at /api/models. Attribute Price per 1M with a link when you republish figures.