Fireworks AI pricing
How Fireworks AI charges, what to know before you commit, and what you'd pay at your usage — next to what the alternatives would cost for the same thing.
How it charges
Checked 2026-10-07 on docs.fireworks.ai ↗ · USD, list pricesIn short
Pay per token for other labs' open-weight models, e.g. $0.15 / $0.60 per million input/output tokens on gpt-oss-120B and $3 / $15 on Kimi K3.
- Free plan
- Yes
- Side project
- $1.25/mo GLM 5.3 Flash
- Growing
- $13/mo GLM 5.3 Flash
- Scaling
- $250/mo GLM 5.3 Flash
What the model leaves out
Serverless standard prices for the models also sold by their own labs. Models not listed are priced by size, from $0.10 per million tokens under 4B parameters.
Before you commit
- Priority mode costs 1.2–1.5× standard, and 'Fast' and US-only variants cost more.
- Batch inference is half price.
| Plan | Monthly fee | Input tokens per month | Output tokens per month |
|---|---|---|---|
| Kimi K3 | None | $3 per million tokens | $15 per million tokens |
| Qwen 3.8 Max | None | $2 per million tokens | $6 per million tokens |
| GLM 5.3 | None | $1.4 per million tokens | $4.4 per million tokens |
| MiniMax M3 | None | $0.3 per million tokens | $1.2 per million tokens |
| DeepSeek V4.1 Flash | None | $0.3 per million tokens | $1.2 per million tokens |
| GLM 5.3 Flash | None | $0.15 per million tokens | $0.5 per million tokens |
| gpt-oss-120B | None | $0.15 per million tokens | $0.6 per million tokens |
What you'd pay
Cost as you grow
All LLM API pricing, compared →Which one fits your situation →