GPU specs chart

Every GPU family you can rent through the providers we track, in one table: memory, bandwidth, tensor throughput at each precision, power and interconnect. Each row was checked against the vendor's own datasheet and links to it. The last column is the cheapest price you can rent that GPU for right now.

Last reviewed . Prices on this page are read live and are not part of the review date.

VRAM at a glance

  • NVIDIA GeForce RTX 5090 VRAM: 32 GB of GDDR7
  • NVIDIA H100 VRAM: 80 to 94 GB of HBM3 (H100 NVL 94 GB, H100 PCIe 80 GB, H100 SXM5 80 GB)
  • NVIDIA H200 VRAM: 141 GB of HBM3e
  • NVIDIA A100 VRAM: 40 to 80 GB of HBM2e (A100 PCIe 40GB 40 GB, A100 PCIe 80GB 80 GB, A100 SXM4 40GB 40 GB, A100 SXM4 80GB 80 GB)
  • NVIDIA B200 VRAM: 180 to 192 GB of HBM3e
  • NVIDIA GeForce RTX 4090 VRAM: 24 GB of GDDR6X
  • NVIDIA L40S VRAM: 48 GB of GDDR6
  • NVIDIA B300 VRAM: 262 to 288 GB of HBM3e
  • NVIDIA RTX PRO 6000 Blackwell VRAM: 96 GB of GDDR7
  • AMD Instinct MI300X VRAM: 192 GB of HBM3

All 60 GPU families

60 of 60 GPU families. Select a column heading to sort.

InterconnectDatasheet
NVIDIA A10Ampere202124 GBGDDR6600125250Not publishedNot publishedNot published150PCIe 4.0$0.37 on LeaderGPUSourcechecked 15 Sep 2026
NVIDIA A100Ampere202040 to 80 GBHBM2e2,039312624Not publishedNot published9.7400NVLink, PCIe 4.0, InfiniBand$0.68 on LeaderGPUSourcechecked 13 Sep 2026
NVIDIA A16Ampere202116 GBGDDR620017.935.9Not publishedNot publishedNot published250PCIe 4.0No current offerSourcechecked 15 Sep 2026
NVIDIA A30Ampere202124 GBHBM2933165330Not publishedNot published5.2165NVLink, PCIe 4.0No current offerSourcechecked 15 Sep 2026
NVIDIA A40Ampere202048 GBGDDR6696149.7299.4Not publishedNot publishedNot published300NVLink, PCIe 4.0$0.49 on RunPodSourcechecked 15 Sep 2026
NVIDIA B200Blackwell2024180 to 192 GBHBM3e8,0002,2504,5004,5009,000371,000NVLink, PCIe 6.0, InfiniBand$3.75 on Packet.aiSourcechecked 13 Sep 2026
NVIDIA B300Blackwell Ultra2025262 to 288 GBHBM3e8,0002,2504,5004,50013,5001.251,400NVLink, PCIe 6.0, InfiniBand$7.89 on RunPodSourcechecked 13 Sep 2026
Intel Gaudi 2Gaudi202296 GBHBM2e2,450Not publishedNot publishedNot publishedNot publishedNot published600RoCE Ethernet, PCIe 4.0$0.91 on LeaderGPUSourcechecked 15 Sep 2026
NVIDIA GB300Blackwell Ultra2025262 to 288 GBHBM3e8,0002,2504,5004,50013,5001.251,400NVLink, PCIe 6.0, InfiniBandNo current offerSourcechecked 13 Sep 2026
NVIDIA GH200 Grace HopperHopper202396 GBHBM34,0009891,9791,979Not published34900NVLink-C2C, PCIe 5.0No current offerSourcechecked 13 Sep 2026
NVIDIA GeForce GTX 1070Pascal20168 GBGDDR5256Not publishedNot publishedNot publishedNot publishedNot published150PCIe 3.0No current offerSourcechecked 15 Sep 2026
NVIDIA GeForce GTX 1080Pascal20168 to 11 GBGDDR5X320Not publishedNot publishedNot publishedNot publishedNot published180PCIe 3.0$0.60 on LeaderGPUSourcechecked 15 Sep 2026
NVIDIA H100Hopper202280 to 94 GBHBM33,3509891,9791,979Not published34700NVLink, PCIe 5.0, InfiniBand$2.50 on HyperstackSourcechecked 13 Sep 2026
NVIDIA H200Hopper2024141 GBHBM3e4,8009891,9791,979Not published34700NVLink, PCIe 5.0, InfiniBand$3.43 on QuantaCloudSourcechecked 13 Sep 2026
NVIDIA L4Ada Lovelace202324 GBGDDR6300121242242Not published0.572PCIe 4.0$0.49 on RunPodSourcechecked 13 Sep 2026
NVIDIA L40Ada Lovelace202348 GBGDDR6864181.05362.1362Not publishedNot published300PCIe 4.0$0.82 on RunPodSourcechecked 15 Sep 2026
NVIDIA L40SAda Lovelace202348 GBGDDR6864362.05733733Not published1.4350PCIe 4.0$0.97 on Massed ComputeSourcechecked 13 Sep 2026
AMD Instinct MI250XCDNA 22021128 GBHBM2e3,277383Not publishedNot publishedNot published47.9560Infinity Fabric, PCIe 4.0No current offerSourcechecked 15 Sep 2026
AMD Instinct MI300XCDNA 32023192 GBHBM35,3001,307.42,614.92,614.9Not published81.7750Infinity Fabric, PCIe 5.0$2.99 on Hot AisleSourcechecked 13 Sep 2026
AMD Instinct MI325XCDNA 32024256 GBHBM3e6,0001,307.42,614.92,614.9Not published81.71,000Infinity Fabric, PCIe 5.0No current offerSourcechecked 15 Sep 2026
AMD Instinct MI355XCDNA 42025288 GBHBM3e8,0002,5005,0005,00010,10078.61,400Infinity Fabric, PCIe 5.0No current offerSourcechecked 15 Sep 2026
NVIDIA Tesla P100Pascal201616 GBHBM273218.7Not publishedNot publishedNot published4.7250NVLink, PCIe 3.0No current offerSourcechecked 15 Sep 2026
NVIDIA Quadro M4000Maxwell20158 GBGDDR5192Not publishedNot publishedNot publishedNot publishedNot published120PCIe 3.0No current offerSourcechecked 15 Sep 2026
NVIDIA Quadro P4000Pascal20178 GBGDDR5243Not publishedNot publishedNot publishedNot publishedNot published105PCIe 3.0$0.51 on PaperspaceSourcechecked 15 Sep 2026
NVIDIA Quadro P5000Pascal201616 GBGDDR5X288Not publishedNot publishedNot publishedNot publishedNot published180PCIe 3.0$0.78 on PaperspaceSourcechecked 15 Sep 2026
NVIDIA Quadro P6000Pascal201624 GBGDDR5X432Not publishedNot publishedNot publishedNot publishedNot published250PCIe 3.0$1.10 on PaperspaceSourcechecked 15 Sep 2026
NVIDIA Quadro RTX 4000Turing20188 GBGDDR641657Not publishedNot publishedNot publishedNot published125PCIe 3.0$0.56 on PaperspaceSourcechecked 15 Sep 2026
NVIDIA Quadro RTX 5000Turing201816 GBGDDR644889.2Not publishedNot publishedNot publishedNot published230NVLink, PCIe 3.0$0.82 on PaperspaceSourcechecked 15 Sep 2026
NVIDIA Quadro RTX 6000Turing201824 GBGDDR6672130.5Not publishedNot publishedNot publishedNot published260NVLink, PCIe 3.0No current offerSourcechecked 15 Sep 2026
NVIDIA Quadro RTX 8000Turing201848 GBGDDR6672130.5Not publishedNot publishedNot publishedNot published260NVLink, PCIe 3.0No current offerSourcechecked 15 Sep 2026
NVIDIA RTX 2000 Ada GenerationAda Lovelace202416 GBGDDR6224Not publishedNot publishedNot publishedNot publishedNot published70PCIe 4.0No current offerSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 2060Turing20196 to 12 GBGDDR6336Not publishedNot publishedNot publishedNot publishedNot published160PCIe 3.0No current offerSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 2070Turing20188 GBGDDR644829.9Not publishedNot publishedNot publishedNot published175PCIe 3.0No current offerSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 2080Turing20188 to 11 GBGDDR644840.3Not publishedNot publishedNot publishedNot published215NVLink, PCIe 3.0No current offerSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 3060Ampere20218 to 12 GBGDDR6360Not publishedNot publishedNot publishedNot publishedNot published170PCIe 4.0$0.10 on Vast.aiSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 3070Ampere20208 GBGDDR644840.681.3Not publishedNot publishedNot published220PCIe 4.0No current offerSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 3080Ampere202010 to 12 GBGDDR6X76059.5119Not publishedNot publishedNot published320PCIe 4.0No current offerSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 3090Ampere202024 GBGDDR6X93671142Not publishedNot publishedNot published350NVLink, PCIe 4.0$0.27 on Vast.aiSourcechecked 15 Sep 2026
NVIDIA RTX 4000 Ada GenerationAda Lovelace202320 GBGDDR6360Not publishedNot publishedNot publishedNot publishedNot published130PCIe 4.0$0.28 on RunPodSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 4060Ada Lovelace20238 to 16 GBGDDR6272Not publishedNot publishedNot publishedNot publishedNot published115PCIe 4.0No current offerSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 4070Ada Lovelace202312 to 16 GBGDDR6X50458.3116.6116.6Not publishedNot published200PCIe 4.0No current offerSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 4080Ada Lovelace202216 GBGDDR6X71797.5195194.9Not publishedNot published320PCIe 4.0No current offerSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 4090Ada Lovelace202224 GBGDDR6X1,008165.2330.4330.3Not published1.3450PCIe 4.0$0.45 on Vast.aiSourcechecked 13 Sep 2026
NVIDIA RTX 4500 AdaAda Lovelace202324 GBGDDR6432Not publishedNot publishedNot publishedNot publishedNot published210PCIe 4.0No current offerSourcechecked 15 Sep 2026
NVIDIA RTX 5000 Ada GenerationAda Lovelace202332 GBGDDR6576Not publishedNot publishedNot publishedNot publishedNot published250PCIe 4.0No current offerSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 5060Blackwell20258 to 16 GBGDDR7448Not publishedNot publishedNot publishedNot publishedNot published145PCIe 5.0$0.18 on Vast.aiSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 5070Blackwell202512 to 16 GBGDDR767261.7123.5123.5493.9Not published250PCIe 5.0No current offerSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 5080Blackwell202516 GBGDDR7960112.6225.1225.1900.4Not published360PCIe 5.0No current offerSourcechecked 15 Sep 2026
NVIDIA GeForce RTX 5090Blackwell202532 GBGDDR71,792209.54194191,6761.6575PCIe 5.0$0.53 on Vast.aiSourcechecked 13 Sep 2026
NVIDIA RTX 5880 AdaAda Lovelace202448 GBGDDR6960Not publishedNot publishedNot publishedNot publishedNot published285PCIe 4.0No current offerSourcechecked 15 Sep 2026
NVIDIA RTX 6000 Ada GenerationAda Lovelace202248 GBGDDR6960364728728.5Not publishedNot published300PCIe 4.0$0.78 on QuantaCloudSourcechecked 15 Sep 2026
NVIDIA RTX A2000Ampere20216 to 12 GBGDDR6288Not published63.9Not publishedNot publishedNot published70PCIe 4.0No current offerSourcechecked 15 Sep 2026
NVIDIA RTX A4000Ampere202116 to 20 GBGDDR6448Not published153.4Not publishedNot publishedNot published140PCIe 4.0$0.15 on HyperstackSourcechecked 15 Sep 2026
NVIDIA RTX A5000Ampere202124 GBGDDR6768Not published222.2Not publishedNot publishedNot published230NVLink, PCIe 4.0$0.23 on Vast.aiSourcechecked 15 Sep 2026
NVIDIA RTX A6000Ampere202048 GBGDDR6768154.8309.6Not publishedNot publishedNot published300NVLink, PCIe 4.0$0.44 on LeaderGPUSourcechecked 15 Sep 2026
NVIDIA RTX PRO 6000 BlackwellBlackwell202596 GBGDDR71,792503.81,007.61,007.62,015.2Not published600PCIe 5.0$1.59 on Vast.aiSourcechecked 15 Sep 2026
NVIDIA Tesla T4Turing201816 GBGDDR632065Not publishedNot publishedNot publishedNot published70PCIe 3.0No current offerSourcechecked 15 Sep 2026
NVIDIA TITAN VVolta201712 GBHBM2653110Not publishedNot publishedNot publishedNot published250PCIe 3.0No current offerSourcechecked 15 Sep 2026
NVIDIA TITAN XpPascal201712 GBGDDR5X548Not publishedNot publishedNot publishedNot publishedNot published250PCIe 3.0No current offerSourcechecked 15 Sep 2026
NVIDIA Tesla V100Volta201716 to 32 GBHBM2900125Not publishedNot publishedNot published7.8300NVLink, PCIe 3.0$0.18 on VERDASourcechecked 13 Sep 2026

TFLOPS columns are tensor throughput as the vendor publishes it. "Not published" means the datasheet gives no figure. VRAM shows a range where a family ships in more than one size. GPU names link to the rent page for that GPU. Prices observed , USD per GPU-hour, on-demand, in stock.

How to read the chart

Dense and with sparsity are different measurements. NVIDIA's datasheets print tensor throughput with a footnote that reads "with sparsity". That figure assumes a model pruned to a 2:4 pattern and is exactly twice what the same GPU does on an ordinary dense model. Almost nobody serves pruned models, so the dense column is the one to compare. For the H100, the datasheet headline is 1,979 TFLOPS at FP16 with sparsity; the dense figure is 989 TFLOPS.

A blank cell is information. FP8 arrived with Ada Lovelace and Hopper, and FP4 with Blackwell, so older GPUs have nothing to publish at those precisions. A model quantised to FP8 still runs on them, but through a slower software path. Consumer GeForce pages leave out FP64, and AMD and Intel publish fewer tensor figures than NVIDIA. We never fill a gap with an estimate.

For LLM inference, read memory first. VRAM decides whether the model fits at all, and memory bandwidth sets the ceiling on tokens per second for a single request, because every generated token reads the weights once. Throughput in TFLOPS matters most for training and for serving many requests at once. To turn a model into a VRAM number, use the LLM VRAM calculator or look the model up in LLM GPU requirements.

TDP is the power budget of the GPU alone. A server adds CPUs, memory, fans and network cards on top. The rent vs buy calculator uses this column to cost the electricity of owning the card.

A short naming decoder

NVIDIA data center parts are named for their architecture: the letter in front is the architecture's initial, and a bigger number within a letter is the bigger or later part. A leading G means the GPU is packaged with a Grace CPU. GeForce and RTX PRO cards count generations in the first digits and the tier in the last two. Suffixes name the form factor or link: PCIe cards fit any server, SXM modules sit on an HGX baseboard with NVLink between all GPUs, and NVL parts are PCIe cards bridged in pairs.

  • Maxwell (NVIDIA, from 2015): Quadro M4000
  • Pascal (NVIDIA, from 2016): GTX 1070, GTX 1080, P100, Quadro P4000, Quadro P5000, Quadro P6000, TITAN Xp
  • Volta (NVIDIA, from 2017): TITAN V, V100
  • Turing (NVIDIA, from 2018): Quadro RTX 4000, Quadro RTX 5000, Quadro RTX 6000, Quadro RTX 8000, RTX 2060, RTX 2070, RTX 2080, T4
  • Ampere (NVIDIA, from 2020): A10, A100, A16, A30, A40, RTX 3060, RTX 3070, RTX 3080, RTX 3090, RTX A2000, RTX A4000, RTX A5000, RTX A6000
  • CDNA 2 (AMD, from 2021): MI250X
  • Ada Lovelace (NVIDIA, from 2022): L4, L40, L40S, RTX 2000 Ada, RTX 4000 Ada, RTX 4060, RTX 4070, RTX 4080, RTX 4090, RTX 4500 Ada, RTX 5000 Ada, RTX 5880 Ada, RTX 6000 Ada
  • Gaudi (Intel, from 2022): Gaudi 2
  • Hopper (NVIDIA, from 2022): GH200, H100, H200
  • CDNA 3 (AMD, from 2023): MI300X, MI325X
  • Blackwell (NVIDIA, from 2024): B200, RTX 5060, RTX 5070, RTX 5080, RTX 5090, RTX PRO 6000
  • Blackwell Ultra (NVIDIA, from 2025): B300, GB300
  • CDNA 4 (AMD, from 2025): MI355X

The full story, generation by generation, is in NVIDIA GPU generations explained.

Method and sources

The figures come from GPUPerHour's spec database, one row per GPU family, served by our public /api/gpu-specs endpoint. Each row was checked by hand against the vendor's datasheet or product page, which is linked in the last column with the date of the check (this release: 15 Sep 2026 and 13 Sep 2026). Where a vendor prints only a with-sparsity figure, the dense figure is that number halved, which is how the vendor defines it. Where a vendor prints nothing, the cell says so.

Prices are not part of the spec data. They are read from live provider listings when the page renders: the lowest on-demand price per GPU-hour among offers that are in stock, from secure providers, seen in the last 15 minutes. MIG slices, fractional vGPU plans and laptop parts are left out. If the price feed is down the column says so instead of showing an old number.

The data is free to reuse under CC BY 4.0; please cite GPUPerHour and link to this page. Found a wrong figure? Tell us through the contact page and we will check it against the datasheet.

Questions

How much VRAM does the H100 have?

The NVIDIA H100 has 80 to 94 GB of HBM3 (H100 NVL 94 GB, H100 PCIe 80 GB, H100 SXM5 80 GB). Its memory bandwidth is 3,350 GB/s.

How much VRAM does the H200 have?

The NVIDIA H200 has 141 GB of HBM3e. Its memory bandwidth is 4,800 GB/s.

How much VRAM does the A100 have?

The NVIDIA A100 has 40 to 80 GB of HBM2e (A100 PCIe 40GB 40 GB, A100 PCIe 80GB 80 GB, A100 SXM4 40GB 40 GB, A100 SXM4 80GB 80 GB). Its memory bandwidth is 2,039 GB/s.

How much VRAM does the RTX 5090 have?

The NVIDIA GeForce RTX 5090 has 32 GB of GDDR7. Its memory bandwidth is 1,792 GB/s.

How much VRAM does the B200 have?

The NVIDIA B200 has 180 to 192 GB of HBM3e. Its memory bandwidth is 8,000 GB/s.

What are the H100's key specs?

NVIDIA H100: Hopper, launched 2022, 80 to 94 GB of HBM3 (H100 NVL 94 GB, H100 PCIe 80 GB, H100 SXM5 80 GB), 3,350 GB/s memory bandwidth, FP16 dense 989 TFLOPS, FP8 dense 1,979 TFLOPS, TDP 700 W.

What is the difference between dense and with-sparsity TFLOPS?

NVIDIA datasheets headline tensor throughput measured with 2:4 structured sparsity, which is exactly twice the dense figure. Most models are not pruned that way, so the dense figure is the one your workload sees. This chart keeps the two in separate columns so you never compare a dense number on one GPU with a sparse number on another.

Why do some cells say Not published?

The vendor's datasheet gives no figure for that precision on that GPU. Older architectures have no FP8 or FP4 hardware, GeForce product pages omit FP64, and AMD and Intel publish fewer tensor figures. We leave the cell blank instead of estimating it or borrowing a figure from another precision.

Where does the cheapest price come from?

It is the lowest current on-demand price per GPU-hour across the providers GPUPerHour tracks, counting only offers that are in stock, from secure (non peer-to-peer) providers, and seen within the last 15 minutes. MIG slices and fractional vGPU plans are excluded because a slice is not the whole GPU the row describes, so the figure can be higher than the lowest price on the rent page.