Fireworks AI vs Together AI
Same workloads, both price lists, refreshed daily. On shared line items today: Fireworks AI is cheaper on 12 of 30 shared models (input or output price), Together AI on 5.
Fireworks AI
Speed-focused open-model inference with FireAttention serving stack, function-calling models and image generation. Strong latency SLAs for production apps.
- Very low latency serving
- Function-calling optimized
- Enterprise SLAs
Together AI
High-performance open-weight inference (Llama, DeepSeek, Qwen) on a custom stack, plus fine-tuning and GPU clusters. Consistently among the fastest and cheapest for open models.
- Top-tier open-model throughput
- Fine-tuning pipeline
- Dedicated endpoints and clusters
Shared models, priced by both
List prices refreshed 2026-10-08 ยท input $/1M tokens| Model | Fireworks AI in/out | Together AI in/out | Cheaper input |
|---|---|---|---|
| GPT-OSS 20B | $0.070 / $0.300 | $0.050 / $0.200 | Together AI |
| DeepSeek-R1-Distill-Qwen-1.5B | $0.100 / $0.100 | $0.180 / $0.180 | Fireworks AI |
| DeepSeek V4 Flash | $0.140 / $0.280 | $0.140 / $0.280 | tie |
| GPT-OSS 120B | $0.150 / $0.600 | $0.150 / $0.600 | tie |
| GLM 5.3 Flash | $0.150 / $0.500 | $0.150 / $0.500 | tie |
| Qwen3 VL 8B Instruct | $0.200 / $0.200 | $0.180 / $0.680 | Together AI |
| Ministral-3-14b-2512 | $0.200 / $0.200 | $0.200 / $0.200 | tie |
| DeepSeek-R1-Distill-Qwen-14B | $0.200 / $0.200 | $1.60 / $1.60 | Fireworks AI |
| NVIDIA-Nemotron-Nano-9B-V2 | $0.200 / $0.200 | $0.060 / $0.250 | Together AI |
| GLM-4.5 Air | $0.220 / $0.880 | $0.200 / $1.10 | Together AI |
| MiniMax M3 | $0.300 / $1.20 | $0.300 / $1.20 | tie |
| DeepSeek V4.1 Flash | $0.300 / $1.20 | $0.300 / $1.20 | tie |
| Muse Glimmer 30B | $0.350 / $1.50 | $0.350 / $1.50 | tie |
| Qwen3.7 Plus | $0.400 / $1.60 | $0.320 / $1.28 | Together AI |
| Qwen3-Coder-480b-A35b-Instruct | $0.450 / $1.80 | $2.00 / $2.00 | Fireworks AI |
| GLM-4.6 | $0.550 / $2.19 | $0.600 / $2.20 | Fireworks AI |
| DeepSeek-V3.1 | $0.560 / $1.68 | $0.600 / $1.70 | Fireworks AI |
| GLM-4.7 | $0.600 / $2.20 | $0.450 / $2.00 | Together AI |
| Kimi K2 | $0.600 / $2.50 | $1.00 / $3.00 | Fireworks AI |
| DeepSeek V3 | $0.900 / $0.900 | $1.25 / $1.25 | Fireworks AI |
| Qwen3 Next 80B A3B Instruct | $0.900 / $0.900 | $0.150 / $1.50 | Together AI |
| Qwen3 Next 80B A3B Thinking | $0.900 / $0.900 | $0.150 / $1.50 | Together AI |
| DeepSeek R1 Distill Llama 70B | $0.900 / $0.900 | $2.00 / $2.00 | Fireworks AI |
| QwQ-32B | $0.900 / $0.900 | $1.20 / $1.20 | Fireworks AI |
| Qwen2p5-Coder-32b | $0.900 / $0.900 | $0.800 / $0.800 | Together AI |
| Qwen3 VL 32B Instruct | $0.900 / $0.900 | $0.500 / $1.50 | Together AI |
| Qwen2.5 VL 72B Instruct | $0.900 / $0.900 | $1.95 / $8.00 | Fireworks AI |
| Qwen2p5-72b | $0.900 / $0.900 | $1.20 / $1.20 | Fireworks AI |
| Mistral Small 3 | $0.900 / $0.900 | $0.100 / $0.300 | Together AI |
| Qwen2-VL-72B-Instruct | $0.900 / $0.900 | $1.20 / $1.20 | Fireworks AI |
Cheaper-on-count is a tally of listed prices, not a quality verdict: throughput, reliability and quantization differ between providers. Disclosure.