OpenRouter vs Together AI
Same workloads, both price lists, refreshed daily. On shared line items today: OpenRouter is cheaper on 20 of 30 shared models (input or output price), Together AI on 1.
OpenRouter
One API key for 340+ models across every major provider, with automatic failover and pass-through pricing. The fastest way to hedge provider risk or A/B models without new integrations.
- 340+ models, one integration
- Automatic provider failover
- Pass-through pricing, small platform fee
Together AI
High-performance open-weight inference (Llama, DeepSeek, Qwen) on a custom stack, plus fine-tuning and GPU clusters. Consistently among the fastest and cheapest for open models.
- Top-tier open-model throughput
- Fine-tuning pipeline
- Dedicated endpoints and clusters
Shared models, priced by both
List prices refreshed 2026-10-08 ยท input $/1M tokens| Model | OpenRouter in/out | Together AI in/out | Cheaper input |
|---|---|---|---|
| GPT-OSS 20B | $0.018 / $0.090 | $0.050 / $0.200 | OpenRouter |
| Llama 3.2 1B Instruct | $0.027 / $0.201 | $0.060 / $0.060 | OpenRouter |
| DeepSeek V4 Flash | $0.030 / $1.28 | $0.140 / $0.280 | OpenRouter |
| GPT-OSS 120B | $0.037 / $0.170 | $0.150 / $0.600 | OpenRouter |
| DeepSeek V4.1 Flash | $0.044 / $0.300 | $0.300 / $1.20 | OpenRouter |
| Llama 3.2 3B | $0.050 / $0.330 | $0.060 / $0.060 | OpenRouter |
| Mistral Small 3 | $0.050 / $0.080 | $0.100 / $0.300 | OpenRouter |
| GLM 5.3 | $0.070 / $7.00 | $1.40 / $4.40 | OpenRouter |
| Gemma 4 31B | $0.090 / $0.340 | $0.390 / $0.970 | OpenRouter |
| Qwen3 Next 80B A3B Instruct | $0.100 / $1.10 | $0.150 / $1.50 | OpenRouter |
| Qwen3.5-9B | $0.100 / $0.150 | $0.170 / $0.250 | OpenRouter |
| Qwen3 VL 32B Instruct | $0.104 / $0.416 | $0.500 / $1.50 | OpenRouter |
| Qwen3 VL 8B Instruct | $0.117 / $0.455 | $0.180 / $0.680 | OpenRouter |
| Qwen3 Coder Next | $0.120 / $0.800 | $0.500 / $1.20 | OpenRouter |
| GLM-4.5 Air | $0.130 / $0.850 | $0.200 / $1.10 | OpenRouter |
| GLM 5.3 Flash | $0.150 / $0.500 | $0.150 / $0.500 | tie |
| Qwen3 Next 80B A3B Thinking | $0.150 / $1.20 | $0.150 / $1.50 | tie |
| Qwen3.8 Flash | $0.150 / $0.470 | $0.150 / $0.470 | tie |
| GLM-5.2 | $0.152 / $12.00 | $1.40 / $4.40 | OpenRouter |
| DeepSeek V4 Pro | $0.209 / $0.418 | $1.32 / $3.96 | OpenRouter |
| MiniMax M2.7 | $0.210 / $0.840 | $0.300 / $1.20 | OpenRouter |
| MiniMax M3 | $0.300 / $1.20 | $0.300 / $1.20 | tie |
| Muse Glimmer 30B | $0.300 / $1.20 | $0.350 / $1.50 | OpenRouter |
| Qwen3.7 Plus | $0.320 / $1.28 | $0.320 / $1.28 | tie |
| GLM-4.6 | $0.430 / $1.75 | $0.600 / $2.20 | OpenRouter |
| Qwen3.5 397B A17B | $0.450 / $3.00 | $0.600 / $3.60 | OpenRouter |
| Nemotron 3 Ultra | $0.500 / $2.20 | $0.600 / $3.60 | OpenRouter |
| Kimi K2 | $0.570 / $2.30 | $1.00 / $3.00 | OpenRouter |
| GLM-4.7 | $0.600 / $2.20 | $0.450 / $2.00 | Together AI |
| GLM-5 | $0.600 / $1.92 | $1.00 / $3.20 | OpenRouter |
Cheaper-on-count is a tally of listed prices, not a quality verdict: throughput, reliability and quantization differ between providers. Disclosure.