Best value cloud GPUs right now

The cheapest GPU per hour is rarely the best value. This page divides each GPU's cheapest live rental price by what you rent it for: gigabytes of memory, dense compute and memory bandwidth. The answer depends on which of those your job runs out of first, so all three are here, recomputed from live prices every hour.

Published 1 October 2026. Prices observed .

The cheapest of each, right now

Per GB of VRAM

Fits the most model in memory for the money. Small cards often win; see the next list for large memory.

  1. NVIDIA A100: 0.85¢ per GB-hour (80 GB at $0.68 per GPU-hour)
  2. NVIDIA GeForce RTX 3060: 0.85¢ per GB-hour (12 GB at $0.10 per GPU-hour)
  3. NVIDIA RTX A6000: 0.92¢ per GB-hour (48 GB at $0.44 per GPU-hour)
  4. NVIDIA RTX A4000: 0.94¢ per GB-hour (16 GB at $0.15 per GPU-hour)
  5. Intel Gaudi 2: 0.95¢ per GB-hour (96 GB at $0.91 per GPU-hour)

Per GB, 48 GB or more

The same measure, limited to sizes that hold a large language model on one card.

  1. NVIDIA A100: 0.85¢ per GB-hour (80 GB at $0.68 per GPU-hour)
  2. NVIDIA RTX A6000: 0.92¢ per GB-hour (48 GB at $0.44 per GPU-hour)
  3. Intel Gaudi 2: 0.95¢ per GB-hour (96 GB at $0.91 per GPU-hour)
  4. NVIDIA A40: 1.02¢ per GB-hour (48 GB at $0.49 per GPU-hour)
  5. NVIDIA RTX 6000 Ada Generation: 1.62¢ per GB-hour (48 GB at $0.78 per GPU-hour)

Per dense FP16 TFLOPS

Most arithmetic for the money: matters for training and for serving many requests at once.

  1. NVIDIA RTX 6000 Ada Generation: 0.21¢ per TFLOPS-hour ($0.78 per GPU-hour)
  2. NVIDIA A100: 0.22¢ per TFLOPS-hour ($0.68 per GPU-hour)
  3. NVIDIA H100: 0.25¢ per TFLOPS-hour ($2.50 per GPU-hour)
  4. NVIDIA GeForce RTX 5090: 0.25¢ per TFLOPS-hour ($0.53 per GPU-hour)
  5. AMD Instinct MI300X: 0.26¢ per TFLOPS-hour ($3.39 per GPU-hour)

Per TB/s of memory bandwidth

Fastest memory for the money: sets the ceiling on tokens per second for a single request.

  1. NVIDIA GeForce RTX 3060: 28.42¢ per TB/s per hour ($0.10 per GPU-hour)
  2. NVIDIA GeForce RTX 5090: 29.76¢ per TB/s per hour ($0.53 per GPU-hour)
  3. NVIDIA GeForce RTX 3090: 30.55¢ per TB/s per hour ($0.29 per GPU-hour)
  4. NVIDIA A100: 33.21¢ per TB/s per hour ($0.68 per GPU-hour)
  5. NVIDIA RTX A4000: 33.48¢ per TB/s per hour ($0.15 per GPU-hour)

All 30 GPU families with a current offer

30 of 30 GPU families with a current offer. Select a column heading to sort.

NVIDIA A10GDDR6$0.37 on LeaderGPU1.55¢ (24 GB)0.30¢Not published61.80¢
NVIDIA A100HBM2e$0.68 on LeaderGPU0.85¢ (80 GB)0.22¢Not published33.21¢
NVIDIA A40GDDR6$0.49 on RunPod1.02¢ (48 GB)0.33¢Not published70.40¢
NVIDIA B200HBM3e$6.79 on RunPod3.54¢ (192 GB)0.30¢0.15¢84.88¢
NVIDIA B300HBM3e$7.89 on RunPod3.01¢ (262 GB)0.35¢0.18¢98.63¢
Intel Gaudi 2HBM2e$0.91 on LeaderGPU0.95¢ (96 GB)Not publishedNot published37.20¢
NVIDIA GeForce GTX 1080GDDR5X$0.60 on LeaderGPU5.45¢ (11 GB)Not publishedNot published$1.88
NVIDIA H100HBM3$2.50 on Hyperstack3.13¢ (80 GB)0.25¢0.13¢74.63¢
NVIDIA H200HBM3e$3.43 on QuantaCloud2.43¢ (141 GB)0.35¢0.17¢71.46¢
NVIDIA L4GDDR6$0.49 on RunPod2.04¢ (24 GB)0.40¢0.20¢$1.63
NVIDIA L40GDDR6$0.79 on ThunderCompute1.65¢ (48 GB)0.44¢0.22¢91.44¢
NVIDIA L40SGDDR6$0.97 on Massed Compute2.02¢ (48 GB)0.27¢0.13¢$1.12
AMD Instinct MI300XHBM3$3.39 on Hot Aisle1.77¢ (192 GB)0.26¢0.13¢63.96¢
NVIDIA Quadro P4000GDDR5$0.51 on Paperspace6.38¢ (8 GB)Not publishedNot published$2.10
NVIDIA Quadro P5000GDDR5X$0.78 on Paperspace4.88¢ (16 GB)Not publishedNot published$2.71
NVIDIA Quadro P6000GDDR5X$1.10 on Paperspace4.58¢ (24 GB)Not publishedNot published$2.55
NVIDIA Quadro RTX 4000GDDR6$0.56 on Paperspace7.00¢ (8 GB)0.98¢Not published$1.35
NVIDIA Quadro RTX 5000GDDR6$0.82 on Paperspace5.13¢ (16 GB)0.92¢Not published$1.83
NVIDIA GeForce RTX 3060GDDR6$0.10 on Vast.ai0.85¢ (12 GB)Not publishedNot published28.42¢
NVIDIA GeForce RTX 3090GDDR6X$0.29 on LeaderGPU1.19¢ (24 GB)0.40¢Not published30.55¢
NVIDIA RTX 4000 Ada GenerationGDDR6$0.28 on RunPod1.40¢ (20 GB)Not publishedNot published77.78¢
NVIDIA GeForce RTX 4090GDDR6X$0.74 on RunPod3.08¢ (24 GB)0.45¢0.22¢73.41¢
NVIDIA GeForce RTX 5060GDDR7$0.18 on Vast.ai1.10¢ (16 GB)Not publishedNot published39.29¢
NVIDIA GeForce RTX 5090GDDR7$0.53 on Vast.ai1.67¢ (32 GB)0.25¢0.13¢29.76¢
NVIDIA RTX 6000 Ada GenerationGDDR6$0.78 on QuantaCloud1.62¢ (48 GB)0.21¢0.11¢80.86¢
NVIDIA RTX A4000GDDR6$0.15 on Hyperstack0.94¢ (16 GB)Not publishedNot published33.48¢
NVIDIA RTX A5000GDDR6$0.27 on RunPod1.13¢ (24 GB)Not publishedNot published35.16¢
NVIDIA RTX A6000GDDR6$0.44 on LeaderGPU0.92¢ (48 GB)0.29¢Not published57.64¢
NVIDIA RTX PRO 6000 BlackwellGDDR7$1.56 on Vast.ai1.62¢ (96 GB)0.31¢0.15¢86.98¢
NVIDIA Tesla V100HBM2$0.83 on Ori2.97¢ (32 GB)0.66¢Not published92.22¢

Costs are USD per hour for one unit: one GB of VRAM, one dense TFLOPS or one TB/s of bandwidth. Lower is better. "Not published" means the vendor gives no figure at that precision, so no value is computed. Prices observed , on-demand, in stock.

Which measure to use

Start with memory. If the model, its KV cache and the runtime do not fit, nothing else matters. Use the per-GB list filtered to the size you need; the LLM VRAM calculator turns a model into a number.

Then ask what runs out first. Generating tokens for one user at a time is limited by memory bandwidth, so the per-TB/s list is the one to read for interactive serving. Training and batched serving are limited by arithmetic, so read the per-TFLOPS list, and the FP8 column if your software runs in FP8. The training versus inference guide and HBM versus GDDR explain why.

When the lists disagree, the job decides. A consumer card can top the per-GB and per-TFLOPS lists while lacking what a production job may need, such as NVLink between cards or a data-center memory system. The RTX PRO 6000 guide compares a professional and a consumer card built on the same chip.

How to read the numbers

Specs are per family. The spec table holds one row per GPU family, so the per-TFLOPS and per-TB/s figures divide the family's cheapest price by the family's published figure. Where a family has slower variants, such as the PCIe version of a GPU sold mainly as SXM, the cheapest offer may be the slower one, and the figure flatters it. The per-GB figure avoids this: it uses the price of the exact memory size. Check the GPU specs chart and the rent page before you book.

Dense figures only. Vendors headline throughput with structured sparsity, which is twice the dense number and applies only to pruned models. This page divides by the dense figure.

Which prices count. The price is the cheapest current on-demand offer per GPU-hour: in stock, from secure (non peer-to-peer) providers, seen in the last 15 minutes. MIG slices, fractional plans and laptop parts are left out, because a slice is not the whole GPU the specs describe. Storage, network transfer and tax are extra.

Method and sources

Specs come from GPUPerHour's spec table, served by the public /api/gpu-specs endpoint, each row checked against the vendor's datasheet. Prices come from live provider listings when the page renders. Price per GB is the cheapest price at each memory size divided by that size, taking the best size in the family. Price per TFLOPS is the family's cheapest price divided by its dense FP16 or FP8 figure. Price per TB/s divides by memory bandwidth in TB/s.

The figures are free to reuse under CC BY 4.0; please cite GPUPerHour and link to this page, with the observation time.

Questions

Which cloud GPU is cheapest per GB of VRAM?

At the latest price observation (1 Oct 2026, 21:55 UTC), the NVIDIA A100: 0.85¢ per GB-hour (80 GB at $0.68 per GPU-hour), from LeaderGPU. Small consumer cards often lead this measure, so check the list filtered to the memory size your model needs.

Which GPU with 48 GB or more is cheapest per GB?

At the latest price observation (1 Oct 2026, 21:55 UTC), the NVIDIA A100: 0.85¢ per GB-hour (80 GB at $0.68 per GPU-hour), from LeaderGPU.

Which cloud GPU is cheapest per TFLOPS?

At the latest price observation (1 Oct 2026, 21:55 UTC), measured on dense FP16 throughput, the NVIDIA RTX 6000 Ada Generation: 0.21¢ per TFLOPS-hour ($0.78 per GPU-hour), from QuantaCloud.

Which cloud GPU is cheapest per unit of memory bandwidth?

At the latest price observation (1 Oct 2026, 21:55 UTC), the NVIDIA GeForce RTX 3060: 28.42¢ per TB/s per hour ($0.10 per GPU-hour), from Vast.ai. Bandwidth sets the ceiling on tokens per second for a single request.

Why use dense rather than with-sparsity TFLOPS?

Vendors headline throughput with 2:4 structured sparsity, which is twice the dense figure and applies only to models pruned that way. An ordinary model runs at the dense rate, so the value figures here divide by the dense number.

How often do these figures change?

The page reads live prices and recomputes at most once an hour. Prices move as providers change rates and stock sells out, so the rankings move too; every figure carries the time of the price observation it used.