Baseten vs OpenRouter

Same workloads, both price lists, refreshed daily. On shared line items today: Baseten is cheaper on 0 of 20 shared models (input or output price), OpenRouter on 12.

Baseten

Production inference platform with Truss packaging, optimized serving engines and enterprise-grade autoscaling. Strong for teams shipping custom models with SLAs.

  • Optimized model serving (TensorRT-LLM)
  • Enterprise autoscaling and observability

Visit Baseten

OpenRouter

One API key for 340+ models across every major provider, with automatic failover and pass-through pricing. The fastest way to hedge provider risk or A/B models without new integrations.

  • 340+ models, one integration
  • Automatic provider failover
  • Pass-through pricing, small platform fee

Visit OpenRouter

Shared models, priced by both

List prices refreshed 2026-10-08 ยท input $/1M tokens
Model Baseten in/out OpenRouter in/out Cheaper input
GPT-OSS 120B $0.100 / $0.500 $0.037 / $0.170 OpenRouter
DeepSeek V4 Flash $0.130 / $0.260 $0.030 / $1.28 OpenRouter
GLM 5.3 Flash $0.150 / $0.500 $0.150 / $0.500 tie
DeepSeek V4.1 Flash $0.300 / $1.20 $0.044 / $0.300 OpenRouter
MiniMax M2.5 $0.300 / $1.20 $0.270 / $1.08 OpenRouter
Inkling Small $0.500 / $1.20 $0.450 / $1.20 OpenRouter
GLM-4.7 $0.600 / $2.20 $0.600 / $2.20 tie
Kimi K2.5 $0.600 / $3.00 $0.450 / $2.25 OpenRouter
Kimi K2 Thinking $0.600 / $2.50 $0.600 / $2.50 tie
Kimi K2 $0.600 / $2.50 $0.570 / $2.30 OpenRouter
GLM-4.6 $0.600 / $2.20 $0.430 / $1.75 OpenRouter
Nemotron 3 Ultra $0.600 / $2.40 $0.500 / $2.20 OpenRouter
Kimi K2.6 $0.950 / $4.00 $0.950 / $4.00 tie
Kimi K2.7 Code $0.950 / $4.00 $0.671 / $3.35 OpenRouter
GLM-5 $0.950 / $3.15 $0.600 / $1.92 OpenRouter
Inkling $1.00 / $4.05 $1.00 / $4.05 tie
GLM-5.2 $1.40 / $4.40 $0.152 / $12.00 OpenRouter
GLM 5.3 $1.40 / $4.40 $0.070 / $7.00 OpenRouter
DeepSeek V4 Pro $1.74 / $3.48 $0.209 / $0.418 OpenRouter
Kimi K3 $3.00 / $15.00 $0.790 / $15.00 OpenRouter

Cheaper-on-count is a tally of listed prices, not a quality verdict: throughput, reliability and quantization differ between providers. Disclosure.