Vicuna API pricing
Ollama · chat model · 2K context · ranked #46 of 1809 by input price
What Vicuna costs in practice
Vicuna is priced at free per million input tokens and free per million output tokens and no separately published prompt-cache rate.
Monthly cost examples
Assuming 1,200 prompt + 400 completion tokens per request unless noted, billed 30 days a month and without cache hits.
| Workload | Cost / month | Cost / day |
|---|---|---|
| 1,000 requests/day | free | free |
| 10,000 requests/day | free | free |
| 100,000 requests/day | free | free |
| 1M summaries/month | free | free |
Model a different profile in the cost calculator.
Capabilities
Cheaper alternatives
Ranked by blended cost (3× input + 1× output) — cheaper, but check context and capabilities before switching.
| # | Model | Provider | Input $/1M | Output $/1M | Cached $/1M | Context |
|---|
Data provenance
This page is assembled from the sources below. Where two sources disagree, both are shown.
Only one source publishes this model; no independent cross-check available yet.
More from Ollama
Frequently asked
› How much does Vicuna cost per 1,000 tokens?
About free for input and free for output tokens, based on the per-million rates on this page.
› Is Vicuna cheaper than other models?
By input price it ranks #46 of 1809 priced chat models in this index, putting it in the cheaper half of the market.
› How often is Vicuna's price checked?
The dataset was last refreshed 2026-09-13 13:40 UTC. Providers can change pricing at any time — verify with Ollama before committing to a budget.