GMKtec EVO-X2 (128GB unified)
The $/GB winner: runs 70B-class and quantized big-MoE (V4 Flash tier) that no consumer GPU fits
Live RTX 5080 quotes (16 GB VRAM) across the providers we track, cheapest first. Marketplace prices are real asks from the Vast.ai order book; secure/on-demand tiers carry datacenter SLAs (what the tiers mean).
Quotes refreshed 2026-08-23| Provider | Tier | $/hr per GPU | Median $/hr | Live offers | |
|---|---|---|---|---|---|
| Vast.ai | marketplace | $0.102 | $0.176 | 64 | Rent |
| RunPod | community | $0.390 | - | - | Rent |
| RunPod | secure | $0.590 | - | - | Rent |
Models from our tracked open-weight set that fit in 16 GB, with KV-cache headroom included. Bigger models need multi-GPU nodes (math in the calculator).
| Model | Params | Fits at | VRAM needed | Cheapest API $/1M out |
|---|---|---|---|---|
| IBM Granite 4.1 8B | 8.8B | FP8 | 13 GB | compare → |
| Llama 3.1 8B | 8B | FP8 | 11 GB | compare → |
| Llama 3.2 3B | 3.2B | FP8 | 5 GB | compare → |
| GPT-OSS 20B | 21B | INT4 | 15 GB | compare → |
| Gemma 3 12B | 12B | INT4 | 9 GB | compare → |
| Phi-4 | 14.7B | INT4 | 11 GB | compare → |
Renting wins for bursty loads; owning wins at sustained 24/7 use. Run your hours through the calculator. If you're buying, these are the machines this audience actually buys:
The $/GB winner: runs 70B-class and quantized big-MoE (V4 Flash tier) that no consumer GPU fits
Bandwidth-per-dollar king for the A3B-MoE and 27-32B dense workhorses
The startup standard: only sub-$10K single card that runs 70B+ comfortably
As an Amazon Associate we earn from qualifying purchases · hardware links may be referral links · picks are editorial, not paid · disclosure
As of 2026-08-23, RTX 5080 rentals start at $0.102/hr (marketplace tier at Vast.ai). Datacenter capacity with SLAs typically costs more than community or marketplace hardware.
With 16 GB of VRAM, a single RTX 5080 fits models up to roughly 11B parameters at FP8 or 22B at INT4 quantization, including KV-cache headroom.
Live badge for a README. It updates itself with every daily refresh:
[](https://llmhosting.ai/gpus/rtx-5080)
Citation: RTX 5080 rental floor $0.102/hr (llmhosting.ai, 2026-08-23). Free to reuse with a link back.