Fireworks AI pricing

How Fireworks AI charges, what to know before you commit, and what you'd pay at your usage — next to what the alternatives would cost for the same thing.

How it charges

Checked 2026-10-07 on docs.fireworks.ai ↗ · USD, list prices
In short

Pay per token for other labs' open-weight models, e.g. $0.15 / $0.60 per million input/output tokens on gpt-oss-120B and $3 / $15 on Kimi K3.

Free plan
Yes
Side project
$1.25/mo GLM 5.3 Flash
Growing
$13/mo GLM 5.3 Flash
Scaling
$250/mo GLM 5.3 Flash
What the model leaves out

Serverless standard prices for the models also sold by their own labs. Models not listed are priced by size, from $0.10 per million tokens under 4B parameters.

Before you commit
  • Priority mode costs 1.2–1.5× standard, and 'Fast' and US-only variants cost more.
  • Batch inference is half price.
PlanMonthly feeInput tokens per monthOutput tokens per month
Kimi K3None$3 per million tokens$15 per million tokens
Qwen 3.8 MaxNone$2 per million tokens$6 per million tokens
GLM 5.3None$1.4 per million tokens$4.4 per million tokens
MiniMax M3None$0.3 per million tokens$1.2 per million tokens
DeepSeek V4.1 FlashNone$0.3 per million tokens$1.2 per million tokens
GLM 5.3 FlashNone$0.15 per million tokens$0.5 per million tokens
gpt-oss-120BNone$0.15 per million tokens$0.6 per million tokens

What you'd pay

Cost as you grow

$0$100$500$1,000$2,0001510501005001k5k
Fireworks AIDeepSeekMistral AIGroqGemini APIClaudeOpenAI APIx: input tokens per month (million tokens), other usage scaled with it · cheapest usable plan at each point, list prices