How this site works
llmhosting.ai is a price terminal for AI compute. Every day, automated pipelines pull prices from provider APIs, cross-check them, and rebuild this site. No hand-edited tables, no "last updated 2024" rot. If the stamp says today, the price was fetched today.
Sources
- LLM list prices: LiteLLM's model price registry (MIT), the de-facto standard machine-readable price file, updated within days of model launches.
- Marketplace quotes: OpenRouter's public models API (transformed and attributed, per their terms).
- Cross-check: Helicone's public cost data (Apache 2.0). Sources disagreeing by more than 50% are excluded from cheapest-price verdicts while remaining visible as that provider's listed price, and 30%+ drift is logged for weekly review.
- GPU rental: live asks from Vast.ai's marketplace API, RunPod's public GraphQL API, and Verda's on-demand and spot list prices.
- Local hardware: live product price and availability from GMKtec's public Shopify product feed; processor and memory specifications from AMD.
What the numbers mean
- Token prices are provider list prices per 1M tokens. Batch discounts, caching and volume deals can take real costs well below list.
- GPU "from" prices are the cheapest live quote per single GPU at fetch time. Marketplace and community tiers are peer hardware: cheaper, fewer guarantees. Secure/on-demand is datacenter capacity.
- Self-hosting figures are planning estimates from parameter counts and rough throughput heuristics, clearly labeled. They are not benchmarks.
Validation
Every refresh passes automated gates before publishing: sanity bounds on every price, minimum row counts (a broken fetch can't wipe the site), and cross-source drift checks. Failures block the publish and page a human.
Funding
Some outbound provider links are referral links; see the full disclosure. Rankings are ordered by price alone; providers cannot pay to move up a table.
LLM prices 2026-08-23 ยท GPU quotes 2026-08-23