GPU comparison

RTX 3090 vs V100

Specifications and current cloud pricing, side by side.

AmperevsVoltaUpdated 15 days ago

Rent the RTX 3090 for single-card LLM inference and fine tuning of models that fit in 24 GB. Its 936 GB/s of GDDR6X edges the V100 at 900 GB/s, its 35.6 TFLOPS of FP32 is more than double the V100 at 15.7 TFLOPS, and as an Ampere part it supports bf16 and current inference frameworks that Volta lacks. Pick the V100 when you need the 32 GB variant for a model that overflows 24 GB, when your code needs its 7.8 TFLOPS of FP64, or when you want a server class card with full NVLink for multi-GPU training, where its 125 TFLOPS of dense FP16 is 1.8x the RTX 3090 at 71 TFLOPS. The RTX 3090 NVLink is a bridge between a pair of consumer cards, not a fabric across a whole server.

RTX 3090 from $0.25/GPU/hrV100 from $0.83/GPU/hr

Right now, from live stock

  • Cheapest right now: RTX 3090 at $0.25/hr on Vast.ai

    Deploy
  • Most providers in stock: RTX 3090 and V100 (2 each)

    See all offers
  • Best $ per TFLOPS FP16 dense right now: RTX 3090 ($0.0035 per TFLOPS-hour at 71 TFLOPS)

    Deploy

Specifications Compared

SpecRTX 3090V100
TDP350W300W
VRAM24 GB16-32 GB
CUDA Cores10,4965,120
Memory TypeGDDR6XHBM2
ArchitectureAmpereVolta
FP16 (dense)71 TFLOPS125 TFLOPS
Form FactorsPCIeSXM2, PCIe
INT8 (dense)284 TOPSNot published
InterconnectNVLink, PCIe 4.0NVLink, PCIe 3.0
Tensor Cores328640
FP32 Performance35.6 TFLOPS15.7 TFLOPS
FP64 PerformanceNot published7.8 TFLOPS
Memory Bandwidth936 GB/s900 GB/s
FP16 (with sparsity)142 TFLOPSNot published
INT8 (with sparsity)568 TOPSNot published

Performance Analysis

The V100 delivers 125 TFLOPS in dense FP16 while the RTX 3090 delivers 71 TFLOPS in dense FP16 so the V100 holds an advantage in dense FP16 workloads. The RTX 3090 supplies 142 TFLOPS in FP16 with sparsity and the V100 supplies no sparsity figure for FP16. Memory bandwidth of 936 GB per second on the RTX 3090 versus 900 GB per second on the V100 allows modestly larger batch sizes on the RTX 3090 during training or inference steps that move large tensors.

Current On-Demand Offers

Cheapest secure on-demand offer per provider that is reported in stock right now, per GPU per hour; multi-GPU instances show the whole-instance price alongside. Deploy opens the DeployGPU console for providers on the platform, otherwise the provider's own site. Stock is rechecked every minute.

RTX 3090

ProviderRegionGPUsPer GPU / hrInstance / hrDeploy
Vast.aiIreland, IE1$0.25—Deploy
LeaderGPUThe Netherlands8$0.29$2.29Deploy

2 providers in stock, 12 offers (cheapest per provider shown). All RTX 3090 offers, price history and alerts

V100

ProviderRegionGPUsPer GPU / hrInstance / hrDeploy
Orilille-12$0.83$1.66Deploy
Paperspaceny21$2.30—Deploy

2 providers in stock, 70 offers (cheapest per provider shown). All V100 offers, price history and alerts

Which GPU to watchWatch the price of

Notify me when RTX 3090 drops below a price

One email when the cheapest in-stock on-demand price per GPU-hour falls below your threshold. Cheapest right now: $0.25/GPU-hr.

QuantaCloud

Comparing providers? We broker across all of them.

Stop tab-switching between pricing pages. Tell us what you need, 16+ GPUs reserved or cluster capacity, and we return one quote at partner rates within 24 hours.

No waitlist24hr quote turnaroundInfiniBand fabric

Compare real-time pricing across 25+ providers

When to Choose the RTX 3090

Choose the RTX 3090 when the whole job fits in 24 GB on one card, which covers most quantized LLM serving, LoRA fine tuning and Stable Diffusion. Its 936 GB/s of bandwidth matches or slightly beats the V100 at 900 GB/s, its 71 TFLOPS of dense FP16 rises to 142 TFLOPS with sparsity, and Ampere supports bf16 while Volta does not. It is a consumer card without ECC memory or datacenter drivers and its NVLink only bridges a pair of cards, so treat it as a single-GPU rental.

When to Choose the V100

The V100 suits workloads that require 125 TFLOPS dense FP16 and that operate within the 300 W TDP envelope. Its NVLink interconnect and availability in SXM2 form factor fit multi GPU scientific computing clusters that rely on dense FP16 throughput.

Use Cases

LLM Training
V100

The V100 supplies 125 TFLOPS dense FP16 which exceeds the 71 TFLOPS dense FP16 of the RTX 3090.

LLM Inference
RTX 3090

The RTX 3090 supplies 35.6 TFLOPS FP32 and sparsity support in INT8 that the V100 does not match.

Fine-tuning
RTX 3090

The RTX 3090 24 GB VRAM and 936 GB per second bandwidth accommodate fine tuning passes better than the lower FP32 rating of the V100.

Stable Diffusion
RTX 3090

The RTX 3090 35.6 TFLOPS FP32 and 24 GB memory suit image generation tasks that rely on FP32 operations.

Scientific Computing
V100

The V100 125 TFLOPS dense FP16 and SXM2 form factor align with established scientific computing clusters.

Frequently Asked Questions

What is the FP32 performance difference?▾

The RTX 3090 reaches 35.6 TFLOPS FP32 and the V100 reaches 15.7 TFLOPS FP32.

How do the memory bandwidth figures compare?▾

The RTX 3090 provides 936 GB per second and the V100 provides 900 GB per second.

What are the TDP ratings?▾

The RTX 3090 has a 350 W TDP and the V100 has a 300 W TDP.

Do both GPUs support NVLink?▾

Both list NVLink, but they are not the same thing. On the V100 in SXM2 form NVLink links every GPU in the server, which is what multi-GPU training relies on. On the RTX 3090 it is a bridge between a pair of cards in a PCIe 4.0 host, so it does not turn a rack of RTX 3090 cards into one training fabric.

Which is cheaper to rent, the RTX 3090 or the V100?▾

Cloud rental prices for both the RTX 3090 and V100 vary by provider, configuration, and availability. This page shows live pricing from 25+ providers updated every 60 seconds. Scroll to the Live Cloud Pricing section to compare current rates.

How much VRAM does the RTX 3090 have compared to the V100?▾

The RTX 3090 has 24 GB of GDDR6X memory. The V100 has 16 to 32 GB of HBM2 memory.

Can I find RTX 3090 and V100 GPUs available to rent right now?▾

Yes. This page shows real-time availability across 25+ cloud GPU providers. The Live Cloud Pricing section displays only in-stock offers with current pricing.

What is the main difference between the RTX 3090 and the V100?▾

The RTX 3090 uses the Ampere architecture (2020) while the V100 uses Volta (2017). The V100 delivers 1.8x the dense FP16 throughput (125 vs 71 TFLOPS, both without sparsity) and the same memory bandwidth as the RTX 3090.

Rent these GPUs

Each GPU page lists every current on-demand offer by provider, per GPU-hour, updated every minute.

Related comparisons

How this page is made

  • Specifications come from the NVIDIA, AMD and Intel datasheets for the RTX 3090 and the V100. Dense and sparse throughput are listed separately.
  • Prices and availability are live: every offer shown is an in-stock, on-demand listing from a provider we track, rechecked every minute.
  • The written comparison (overview, performance notes, when to choose each card, verdict, use cases and FAQ) was drafted with an AI model from the specification table above and passed an automated check that rejects any figure not in that table. It was last generated on . It contains no prices; those are always read live.
  • Read how we collect the data, or report an error on this page.

Next steps