Qwen3 Embedding 4b API pricing
Fireworks AI · embedding model · 41K context
What Qwen3 Embedding 4b costs in practice
Qwen3 Embedding 4b is priced at free per million input tokens and free per million output tokens and no separately published prompt-cache rate.
Monthly cost examples
Assuming 1,200 prompt + 400 completion tokens per request unless noted, billed 30 days a month and without cache hits.
| Workload | Cost / month | Cost / day |
|---|---|---|
| 1,000 requests/day | free | free |
| 10,000 requests/day | free | free |
| 100,000 requests/day | free | free |
| 1M summaries/month | free | free |
Model a different profile in the cost calculator.
Capabilities
Cheaper alternatives
Ranked by blended cost (3× input + 1× output) — cheaper, but check context and capabilities before switching.
| # | Model | Provider | Input $/1M | Output $/1M | Cached $/1M | Context |
|---|
Data provenance
This page is assembled from the sources below. Where two sources disagree, both are shown.
Only one source publishes this model; no independent cross-check available yet.
More from Fireworks AI
Frequently asked
› How much does Qwen3 Embedding 4b cost per 1,000 tokens?
About free for input and free for output tokens, based on the per-million rates on this page.
› Is Qwen3 Embedding 4b cheaper than other models?
We do not have enough price data to rank this model.
› How often is Qwen3 Embedding 4b's price checked?
The dataset was last refreshed 2026-09-13 13:40 UTC. Providers can change pricing at any time — verify with Fireworks AI before committing to a budget.