H200 vs MI300X

HoppervsCDNA 3Updated today

The MI300X wins for the most common use case of LLM training because it provides 1307.4 TFLOPS dense FP16 performance and 192 GB memory capacity which exceed the H200 specifications in those critical metrics.

H200 from $3.43/GPU/hrMI300X from $2.39/GPU/hr

Right now, from live stock

  • Cheapest right now: MI300X at $2.39/hr on RunPod

    Deploy
  • Most providers in stock: H200 (8)

    See all offers
  • Best $ per TFLOPS FP16 dense right now: MI300X ($0.0018 per TFLOPS-hour at 1,307.4 TFLOPS)

    Deploy

Specifications Compared

SpecH200MI300X
TDP700W750W
VRAM141 GB192 GB
CUDA Cores16,896Not published
FP8 (dense)1,979 TFLOPS2,614.9 TFLOPS
Memory TypeHBM3eHBM3
ArchitectureHopperCDNA 3
FP16 (dense)989 TFLOPS1,307.4 TFLOPS
Form FactorsSXM, NVLOAM
INT8 (dense)1,979 TOPS2,614.9 TOPS
InterconnectNVLink, PCIe 5.0, InfiniBandInfinity Fabric, PCIe 5.0
Tensor Cores528Not published
FP32 Performance67 TFLOPS163.4 TFLOPS
FP64 Performance34 TFLOPS81.7 TFLOPS
Memory Bandwidth4,800 GB/s5,300 GB/s
FP8 (with sparsity)3,958 TFLOPS5,229.8 TFLOPS
FP16 (with sparsity)1,979 TFLOPS2,614.9 TFLOPS
INT8 (with sparsity)3,958 TOPS5,229.8 TOPS

Performance Analysis

The dense FP16 performance positions the MI300X at 1307.4 TFLOPS against the H200 at 989 TFLOPS which translates to greater throughput for training large models where dense operations dominate. The FP32 performance gap shows the MI300X at 163.4 TFLOPS versus the H200 at 67 TFLOPS indicating stronger suitability for scientific simulations that rely on full precision calculations. Memory bandwidth differences mean the MI300X at 5300 GB/s supports larger batch sizes during inference compared to the H200 at 4800 GB/s while the H200 TDP of 700W remains lower than the MI300X TDP of 750W.

Current On-Demand Offers

Cheapest secure on-demand offer per provider that is reported in stock right now, per GPU per hour; multi-GPU instances show the whole-instance price alongside. Deploy opens the DeployGPU console for providers on the platform, otherwise the provider's own site. Stock is rechecked every minute.

H200

ProviderRegionGPUsPer GPU / hrInstance / hrDeploy
QuantaCloudus-east-11$3.43Deploy
Orilondon-31$3.50Deploy
Massed Computeus-east-14$3.62$14.48Deploy
VERDAFIN-HEL1$4.24Deploy
LyceumEurope4$4.29$17.16Deploy

8 providers in stock, 15 offers (cheapest per provider shown). All H200 offers, price history and alerts

MI300X

ProviderRegionGPUsPer GPU / hrInstance / hrDeploy
RunPodglobal1$2.39Deploy
Hot AisleMichigan1$2.99Deploy

2 providers in stock, 3 offers (cheapest per provider shown). All MI300X offers, price history and alerts

Which GPU to watchWatch the price of

Notify me when H200 drops below a price

One email when the cheapest in-stock on-demand price per GPU-hour falls below your threshold. Cheapest right now: $3.43/GPU-hr.

QuantaCloud

Comparing H-series providers? We broker across all of them.

Hopper stock changes by the hour and prices differ by provider. If you need 16+ GPUs reserved or a cluster in the next 90 days, we quote H-series or B300 inventory at partner rates: one quote, 24h turnaround.

No waitlist24hr quote turnaroundInfiniBand fabric

Compare real-time pricing across 25+ providers

When to Choose the H200

The H200 suits deployments that prioritize lower power consumption at 700W TDP and NVLink interconnect for multi GPU scaling in dense clusters. Organizations running workloads sensitive to energy efficiency select the H200 when form factors include SXM or NVL options.

When to Choose the MI300X

The MI300X fits scenarios requiring maximum memory at 192 GB and bandwidth at 5300 GB/s for handling very large models or datasets. Users focused on higher dense FP16 throughput at 1307.4 TFLOPS choose the MI300X for training and inference tasks that benefit from CDNA 3 architecture.

Use Cases

LLM Training
MI300X

The MI300X supplies 1307.4 TFLOPS dense FP16 performance and 192 GB memory which exceed H200 figures for large scale training.

LLM Inference
MI300X

The MI300X offers 192 GB memory and 5300 GB/s bandwidth which enable larger batch sizes than the H200 configuration.

Fine-tuning
MI300X

The MI300X delivers higher dense FP16 at 1307.4 TFLOPS and greater memory capacity for efficient parameter updates.

Stable Diffusion
Either

Both GPUs handle the workload adequately with the H200 at 989 TFLOPS dense FP16 and the MI300X at 1307.4 TFLOPS dense FP16.

Scientific Computing
MI300X

The MI300X provides 163.4 TFLOPS FP32 performance which surpasses the H200 at 67 TFLOPS for precision sensitive computations.

Frequently Asked Questions

How does FP16 performance compare between these GPUs?

The H200 achieves 989 TFLOPS dense FP16 while the MI300X achieves 1307.4 TFLOPS dense FP16. The MI300X also reaches 2614.9 TFLOPS with sparsity compared to the H200 at 1979 TFLOPS with sparsity.

What interconnect options exist for the H200 and MI300X?

The H200 supports NVLink along with PCIe 5.0 and InfiniBand. The MI300X supports Infinity Fabric along with PCIe 5.0.

Which GPU has higher memory bandwidth?

The MI300X provides 5300 GB/s bandwidth compared to the H200 at 4800 GB/s. This difference affects data movement rates during model execution.

What are the TDP ratings for H200 and MI300X?

The H200 lists a TDP of 700W while the MI300X lists a TDP of 750W. Power requirements influence cooling and operational costs in large installations.

How does FP32 performance differ between the two?

The H200 delivers 67 TFLOPS FP32 while the MI300X delivers 163.4 TFLOPS FP32. The MI300X advantage benefits applications that require high precision arithmetic.

Which is cheaper to rent, the H200 or the MI300X?

Cloud rental prices for both the H200 and MI300X vary by provider, configuration, and availability. This page shows live pricing from 25+ providers updated every 60 seconds. Scroll to the Live Cloud Pricing section to compare current rates.

How much VRAM does the H200 have compared to the MI300X?

The H200 has 141 GB of HBM3e memory. The MI300X has 192 GB of HBM3 memory.

Can I find H200 and MI300X GPUs available to rent right now?

Yes. This page shows real-time availability across 25+ cloud GPU providers. The Live Cloud Pricing section displays only in-stock offers with current pricing.

What is the main difference between the H200 and the MI300X?

The H200 uses the Hopper architecture (2024) while the MI300X uses CDNA 3 (2023). The MI300X delivers 1.3x the dense FP16 throughput (1,307.4 vs 989 TFLOPS, both without sparsity) and 1.1x the memory bandwidth of the H200.

Rent these GPUs

Each GPU page lists every current on-demand offer by provider, per GPU-hour, updated every minute.

Related comparisons

How this page is made

  • Specifications come from the NVIDIA, AMD and Intel datasheets for the H200 and the MI300X. Dense and sparse throughput are listed separately.
  • Prices and availability are live: every offer shown is an in-stock, on-demand listing from a provider we track, rechecked every minute.
  • The written comparison (overview, performance notes, when to choose each card, verdict, use cases and FAQ) was drafted with an AI model from the specification table above and passed an automated check that rejects any figure not in that table. It was last generated on . It contains no prices; those are always read live.
  • Read how we collect the data, or report an error on this page.

Next steps