DeepInfra vs Nebius

Same workloads, both price lists, refreshed daily. On shared line items today: DeepInfra is cheaper on 11 of 16 shared models (input or output price), Nebius on 2.

DeepInfra

Consistently the price floor for open-weight models: Llama, Qwen, DeepSeek at rock-bottom per-token rates with an OpenAI-compatible API and per-request GPU billing options..

  • Among the cheapest open-model tokens anywhere
  • OpenAI-compatible API
  • Dedicated GPU deployments

Visit DeepInfra

Nebius

AI cloud from the former Yandex team with H100/H200/B200 clusters in Europe and the US, plus Nebius AI Studio for per-token open-weight inference. Competitive pricing at cluster scale.

  • European datacenters
  • Both clusters and per-token inference
  • Aggressive large-scale pricing

Visit Nebius

Shared models, priced by both

List prices refreshed 2026-08-23 ยท input $/1M tokens
Model DeepInfra in/out Nebius in/out Cheaper input
Mistral Nemo $0.020 / $0.040 $0.040 / $0.120 DeepInfra
Llama 3.1 8B $0.030 / $0.050 $0.020 / $0.060 Nebius
Llama-Guard-3-8B $0.055 / $0.055 $0.020 / $0.060 Nebius
Qwen3 14B $0.060 / $0.240 $0.080 / $0.240 DeepInfra
Qwen3 30B A3B $0.080 / $0.290 $0.100 / $0.300 DeepInfra
Gemma 3 27B $0.090 / $0.160 $0.060 / $0.200 Nebius
Qwen3 32B $0.100 / $0.280 $0.100 / $0.300 tie
Qwen2p5-72b $0.120 / $0.390 $0.130 / $0.400 DeepInfra
QwQ-32B $0.150 / $0.400 $0.150 / $0.450 tie
Qwen3 235B A22B $0.180 / $0.540 $0.200 / $0.600 DeepInfra
DeepSeek R1 Distill Llama 70B $0.200 / $0.600 $0.250 / $0.750 DeepInfra
Llama 3.3 70B $0.230 / $0.400 $0.130 / $0.400 Nebius
DeepSeek V3 $0.380 / $0.890 $0.500 / $1.50 DeepInfra
Llama 3.1 70B Instruct $0.400 / $0.400 $0.130 / $0.400 Nebius
DeepSeek R1 $0.700 / $2.40 $0.800 / $2.40 DeepInfra
Hermes 3 405B Instruct $1.00 / $1.00 $1.00 / $3.00 tie

Cheaper-on-count is a tally of listed prices, not a quality verdict: throughput, reliability and quantization differ between providers. Disclosure.