Fireworks AI vs OpenRouter
Same workloads, both price lists, refreshed daily. On shared line items today: Fireworks AI is cheaper on 5 of 30 shared models (input or output price), OpenRouter on 15.
Fireworks AI
Speed-focused open-model inference with FireAttention serving stack, function-calling models and image generation. Strong latency SLAs for production apps.
- Very low latency serving
- Function-calling optimized
- Enterprise SLAs
OpenRouter
One API key for 340+ models across every major provider, with automatic failover and pass-through pricing. The fastest way to hedge provider risk or A/B models without new integrations.
- 340+ models, one integration
- Automatic provider failover
- Pass-through pricing, small platform fee
Shared models, priced by both
List prices refreshed 2026-10-08 ยท input $/1M tokens| Model | Fireworks AI in/out | OpenRouter in/out | Cheaper input |
|---|---|---|---|
| GPT-OSS 20B | $0.070 / $0.300 | $0.018 / $0.090 | OpenRouter |
| DeepSeek V4 Flash | $0.140 / $0.280 | $0.030 / $1.28 | OpenRouter |
| GPT-OSS 120B | $0.150 / $0.600 | $0.037 / $0.170 | OpenRouter |
| GLM 5.3 Flash | $0.150 / $0.500 | $0.150 / $0.500 | tie |
| Qwen3 30B A3B | $0.150 / $0.600 | $0.120 / $0.500 | OpenRouter |
| Qwen3 Coder 30B A3B Instruct | $0.150 / $0.600 | $0.070 / $0.280 | OpenRouter |
| Qwen3 VL 30B A3B Instruct | $0.150 / $0.600 | $0.150 / $0.600 | tie |
| Qwen3 VL 30B A3B Thinking | $0.150 / $0.600 | $0.200 / $2.40 | Fireworks AI |
| Mistral Nemo | $0.200 / $0.200 | $0.019 / $0.030 | OpenRouter |
| Qwen3 14B | $0.200 / $0.200 | $0.120 / $0.240 | OpenRouter |
| Qwen3 VL 8B Instruct | $0.200 / $0.200 | $0.117 / $0.455 | OpenRouter |
| Mistral-7b | $0.200 / $0.200 | $0.130 / $0.130 | OpenRouter |
| MythoMax 13B | $0.200 / $0.200 | $0.080 / $0.110 | OpenRouter |
| Qwen3 8B | $0.200 / $0.200 | $0.117 / $0.455 | OpenRouter |
| Qwen3 235B A22B | $0.220 / $0.880 | $0.455 / $1.82 | Fireworks AI |
| Qwen3 VL 235B A22B Instruct | $0.220 / $0.880 | $0.210 / $1.90 | OpenRouter |
| GLM-4.5 Air | $0.220 / $0.880 | $0.130 / $0.850 | OpenRouter |
| Qwen3 VL 235B A22B Thinking | $0.220 / $0.880 | $0.400 / $4.00 | Fireworks AI |
| Qwen3 235B A22B Thinking 2507 | $0.220 / $0.880 | $0.230 / $2.30 | Fireworks AI |
| DeepSeek V4 Flash Vision Exp | $0.220 / $0.660 | $0.216 / $0.647 | OpenRouter |
| MiniMax M3 | $0.300 / $1.20 | $0.300 / $1.20 | tie |
| DeepSeek V4.1 Flash | $0.300 / $1.20 | $0.044 / $0.300 | OpenRouter |
| MiniMax M2.1 | $0.300 / $1.20 | $0.300 / $1.20 | tie |
| MiniMax M2 | $0.300 / $1.20 | $0.300 / $1.20 | tie |
| Muse Glimmer 30B | $0.350 / $1.50 | $0.300 / $1.20 | OpenRouter |
| Qwen3.7 Plus | $0.400 / $1.60 | $0.320 / $1.28 | OpenRouter |
| GPT-OSS-Safeguard-20b | $0.500 / $0.500 | $0.075 / $0.300 | OpenRouter |
| GLM-4.6 | $0.550 / $2.19 | $0.430 / $1.75 | OpenRouter |
| GLM-4.5 | $0.550 / $2.19 | $0.600 / $2.20 | Fireworks AI |
| DeepSeek V3.2 | $0.560 / $1.68 | $0.280 / $0.420 | OpenRouter |
Cheaper-on-count is a tally of listed prices, not a quality verdict: throughput, reliability and quantization differ between providers. Disclosure.