Price per 1M logo

Cheapest model for a customer-support bot

Blended cost estimate for a support bot profile: short prompts, medium answers, high volume, heavy prompt reuse.

Current leader: Gemini Exp 1114 (Google Gemini). Refreshed 2026-09-13 13:40 UTC.

#ModelProviderInput $/1MOutput $/1MCached $/1MContext
1Gemini Exp 1114Google Geminifreefree1.0M
2Gemini Exp 1206Google Geminifreefree2.1M
3Gemma 3 27b ItGoogle Geminifreefree131K
4Gemma 4 26b A4b ItGoogle Geminifreefree262K
5Gemma 4 31b ItGoogle Geminifreefree262K
6Learnlm 1.5 Pro ExperimentalGoogle Geminifreefree33K
7Labs Leanstral 1 5Mistral AIfreefree262K
8Labs Leanstral 1 5 1Mistral AIfreefree262K
9Anthropic.claude Mythos PreviewAmazon Bedrockfreefree1M
10Gemma 2b It LoraCloudflare Workers AIfreefree8K
11Mistral 7b Instruct V0.2 LoraCloudflare Workers AIfreefree15K
12Llama 2 7b Chat Hf LoraCloudflare Workers AIfreefree8K
13Gemma 7b It LoraCloudflare Workers AIfreefree4K
14Codegeex4Ollamafreefree33K
15CodegemmaOllamafreefree8K
16CodellamaOllamafreefree4K
17Deepseek Coder V2 BaseOllamafreefree8K
18Deepseek Coder V2 InstructOllamafreefree33K
19Deepseek Coder V2 Lite BaseOllamafreefree8K
20Deepseek Coder V2 Lite InstructOllamafreefree33K
21Deepseek V3.1:671b CloudOllamafreefree164K
22GPT Oss:120b CloudOllamafreefree131K
23GPT Oss:20b CloudOllamafreefree131K
24Internlm2 5 20b ChatOllamafreefree33K
25Llama2Ollamafreefree4K
26Llama2 UncensoredOllamafreefree4K
27Llama2:13bOllamafreefree4K
28Llama2:70bOllamafreefree4K
29Llama2:7bOllamafreefree4K
30Llama3Ollamafreefree8K
31Llama3.1Ollamafreefree8K
32Llama3:70bOllamafreefree8K
33Llama3:8bOllamafreefree8K
34MistralOllamafreefree8K
35Mistral 7B Instruct V0.1Ollamafreefree8K
36Mistral 7B Instruct V0.2Ollamafreefree33K
37Mistral Large Instruct 2407Ollamafreefree66K
38Mixtral 8x22B Instruct V0.1Ollamafreefree66K
39Mixtral 8x7B Instruct V0.1Ollamafreefree33K
40Orca MiniOllamafreefree4K
41Qwen3 Coder:480b CloudOllamafreefree262K
42VicunaOllamafreefree2K
43AutoOpenRouterfreefree2M
44FreeOpenRouterfreefree200K
45BodybuilderOpenRouterfreefree128K
46Nemotron 3.5 Lightning:freeOpenRouterfreefree1M
47Laguna S 2.1:freeOpenRouterfreefree262K
48Laguna Xs 2.1:freeOpenRouterfreefree262K
49Glm 5.2:freeOpenRouterfreefree256K
50Nemotron 3.5 Content Safety:freeOpenRouterfreefree128K
51Nemotron 3 Ultra 550b A55b:freeOpenRouterfreefree1M
52Minimax M3:freeOpenRouterfreefree1.0M
53Nemotron 3 Nano Omni 30b A3b Reasoning:freeOpenRouterfreefree256K
54Gemma 4 26b A4b It:freeOpenRouterfreefree262K
55Gemma 4 31b It:freeOpenRouterfreefree262K
56Minimax M2.7:freeOpenRouterfreefree197K
57Nemotron 3 Super 120b A12b:freeOpenRouterfreefree262K
58Llama 3.3 70B Instruct Turbo FreeTogether AIfreefree
59Ternary Bonsai 27BTogether AIfreefree262K
60Qwen2.5 Coder 7BNebius$0.01$0.0333K

Other use cases

Frequently asked

Which AI model should a support chatbot use?

Gemini Exp 1114 from Google Gemini currently leads this list, at $0/1M input tokens. Rankings are recomputed from the raw dataset on every refresh, so they track price cuts automatically.

How are these rankings calculated?

Each use case applies a capability filter (context size, vision, tool calling, cache pricing) and then ranks the surviving models by the blended cost that profile actually pays. See the methodology page for the exact formulas.

Is the cheapest model always the right choice?

No. Check the capability flags and context window, and benchmark quality on your own data — the cheapest model that fails your evals costs more than the second cheapest that passes them.