GPU comparison

P100 vs V100

Specifications and current cloud pricing, side by side.

PascalvsVoltaUpdated 15 days ago

Rent the V100. It has tensor cores that the P100 lacks, which is why dense FP16 is 125 TFLOPS against 18.7 TFLOPS, a 6.7x gap, and it adds 900 GB/s of bandwidth against 732 GB/s and a 32 GB option against a fixed 16 GB. The only condition that favors the P100 is a pure FP32 or FP64 batch job that fits in 16 GB, is not time sensitive, and where the P100's 9.3 TFLOPS FP32 or 4.7 TFLOPS FP64 is enough. For anything involving mixed precision, transformers or diffusion models, the P100 is too slow to be worth renting.

P100 listed from $0.09/GPU/hrV100 listed from $0.83/GPU/hr

Right now, from live stock

Specifications Compared

SpecP100V100
TDP250W300W
VRAM16 GB16-32 GB
CUDA Cores3,5845,120
Memory TypeHBM2HBM2
ArchitecturePascalVolta
FP16 (dense)18.7 TFLOPS125 TFLOPS
Form FactorsSXM2, PCIeSXM2, PCIe
InterconnectNVLink, PCIe 3.0NVLink, PCIe 3.0
Tensor CoresNot published640
FP32 Performance9.3 TFLOPS15.7 TFLOPS
FP64 Performance4.7 TFLOPS7.8 TFLOPS
Memory Bandwidth732 GB/s900 GB/s

Performance Analysis

The dense FP16 rating reaches 18.7 TFLOPS on the P100 and 125 TFLOPS on the V100. This gap indicates that workloads using dense FP16 operations complete faster on the V100. Memory bandwidth at 900 GB/s on the V100 compared with 732 GB/s on the P100 supports larger batch sizes during training sessions that rely on dense FP16 arithmetic. FP32 throughput at 15.7 TFLOPS versus 9.3 TFLOPS further favors the V100 for mixed precision pipelines that alternate between dense FP32 and dense FP16 stages.

Current On-Demand Offers

Cheapest secure on-demand offer per provider that is reported in stock right now, per GPU per hour; multi-GPU instances show the whole-instance price alongside. Deploy opens the DeployGPU console for providers on the platform, otherwise the provider's own site. Stock is rechecked every minute.

P100

P100 is not offered on-demand by any provider we track right now. See the P100 rental page for last-seen listed prices and a price alert.

V100

ProviderRegionGPUsPer GPU / hrInstance / hrDeploy
Orilille-12$0.83$1.66Deploy
Paperspaceny21$2.30—Deploy

2 providers in stock, 70 offers (cheapest per provider shown). All V100 offers, price history and alerts

Which GPU to watchWatch the price of

Notify me when P100 drops below a price

One email when the cheapest in-stock on-demand price per GPU-hour falls below your threshold.

QuantaCloud

Comparing providers? We broker across all of them.

Stop tab-switching between pricing pages. Tell us what you need, 16+ GPUs reserved or cluster capacity, and we return one quote at partner rates within 24 hours.

No waitlist24hr quote turnaroundInfiniBand fabric

Compare real-time pricing across 25+ providers

When to Choose the P100

Choose the P100 only for FP32 or FP64 work that fits in 16 GB and does not use tensor cores: it lists 9.3 TFLOPS FP32 and 4.7 TFLOPS FP64, both roughly 60 percent of the V100 figures, and 732 GB/s of HBM2 bandwidth. It also has NVLink and the same SXM2 and PCIe form factors, so multi card FP64 simulations are possible. Do not choose it for LLM inference, fine tuning or image generation; with no published tensor core count, its 18.7 TFLOPS of FP16 is about 15 percent of the V100 figure.

When to Choose the V100

The V100 suits training sessions that exploit dense FP16 throughput of 125 TFLOPS and memory bandwidth of 900 GB/s. Workloads that scale with 16 to 32 GB HBM2 configurations benefit from the higher FP32 rating of 15.7 TFLOPS at 300W TDP.

Use Cases

LLM Training
V100

The V100 supplies dense FP16 performance of 125 TFLOPS against 18.7 TFLOPS on the P100.

LLM Inference
V100

Higher dense FP16 throughput of 125 TFLOPS on the V100 reduces latency for inference batches.

Fine-tuning
V100

Memory bandwidth of 900 GB/s on the V100 accommodates larger fine tuning batches than 732 GB/s on the P100.

Stable Diffusion
Either

Both GPUs provide 16 GB HBM2 and support the required form factors for diffusion model execution.

Scientific Computing
P100

The P100 meets requirements at a TDP of 250W when dense FP16 acceleration is not essential.

Frequently Asked Questions

What is the dense FP16 performance difference between the P100 and V100?▾

The P100 lists dense FP16 at 18.7 TFLOPS while the V100 lists dense FP16 at 125 TFLOPS.

How does memory bandwidth compare on the P100 versus the V100?▾

The P100 provides 732 GB/s bandwidth and the V100 provides 900 GB/s bandwidth.

What FP32 performance do the P100 and V100 deliver?▾

The P100 reaches 9.3 TFLOPS in FP32 and the V100 reaches 15.7 TFLOPS in FP32.

Which GPU has higher TDP between the P100 and V100?▾

The P100 operates at 250W TDP and the V100 operates at 300W TDP.

Do the P100 and V100 share the same interconnect options?▾

Both GPUs support NVLink and PCIe 3.0 interconnects along with SXM2 and PCIe form factors.

Which is cheaper to rent, the P100 or the V100?▾

Cloud rental prices for both the P100 and V100 vary by provider, configuration, and availability. This page shows live pricing from 25+ providers updated every 60 seconds. Scroll to the Live Cloud Pricing section to compare current rates.

How much VRAM does the P100 have compared to the V100?▾

The P100 has 16 GB of HBM2 memory. The V100 has 16 to 32 GB of HBM2 memory.

Can I find P100 and V100 GPUs available to rent right now?▾

Yes. This page shows real-time availability across 25+ cloud GPU providers. The Live Cloud Pricing section displays only in-stock offers with current pricing.

What is the main difference between the P100 and the V100?▾

The P100 uses the Pascal architecture (2016) while the V100 uses Volta (2017). The V100 delivers 6.7x the dense FP16 throughput (125 vs 18.7 TFLOPS, both without sparsity) and 1.2x the memory bandwidth of the P100.

Rent these GPUs

Each GPU page lists every current on-demand offer by provider, per GPU-hour, updated every minute.

Related comparisons

How this page is made

  • Specifications come from the NVIDIA, AMD and Intel datasheets for the P100 and the V100. Dense and sparse throughput are listed separately.
  • Prices and availability are live: every offer shown is an in-stock, on-demand listing from a provider we track, rechecked every minute.
  • The written comparison (overview, performance notes, when to choose each card, verdict, use cases and FAQ) was drafted with an AI model from the specification table above and passed an automated check that rejects any figure not in that table. It was last generated on . It contains no prices; those are always read live.
  • Read how we collect the data, or report an error on this page.

Next steps