Specifications Compared
| Spec | A16 | V100 |
|---|---|---|
| TDP | 250W | 300W |
| VRAM | 16 GB | 16-32 GB |
| CUDA Cores | 1,280 | 5,120 |
| Memory Type | GDDR6 | HBM2 |
| Architecture | Ampere | Volta |
| FP16 (dense) | 17.9 TFLOPS | 125 TFLOPS |
| Form Factors | PCIe | SXM2, PCIe |
| INT8 (dense) | 35.9 TOPS | Not published |
| Interconnect | PCIe 4.0 | NVLink, PCIe 3.0 |
| Tensor Cores | 40 | 640 |
| FP32 Performance | 4.5 TFLOPS | 15.7 TFLOPS |
| FP64 Performance | Not published | 7.8 TFLOPS |
| Memory Bandwidth | 200 GB/s | 900 GB/s |
| FP16 (with sparsity) | 35.9 TFLOPS | Not published |
| INT8 (with sparsity) | 71.8 TOPS | Not published |
Performance Analysis
The FP16 dense performance shows the V100 at 125 TFLOPS exceeding the A16 at 17.9 TFLOPS. This delta implies faster training and inference times for models using dense FP16 operations on the V100. The FP32 figures of 4.5 TFLOPS on the A16 and 15.7 TFLOPS on the V100 follow a similar pattern for general computations. Memory bandwidth at 200 GB/s on the A16 compared to 900 GB/s on the V100 limits batch sizes in memory intensive tasks for the A16. Sparse FP16 performance reaches 35.9 TFLOPS on the A16 but remains not published for the V100 so direct comparison applies only to dense figures. The A16 provides INT8 dense at 35.9 TOPS which supports quantized inference workloads not specified for the V100.
Current On-Demand Offers
Cheapest secure on-demand offer per provider that is reported in stock right now, per GPU per hour; multi-GPU instances show the whole-instance price alongside. Deploy opens the DeployGPU console for providers on the platform, otherwise the provider's own site. Stock is rechecked every minute.
A16
| Provider | Region | GPUs | Per GPU / hr | Instance / hr | Deploy |
|---|---|---|---|---|---|
| Vultr | Silicon Valley, US | 1 | $0.06 | — | Deploy |
| Ori | bangalore | 4 | $0.50 | $2.00 | Deploy |
2 providers in stock, 22 offers (cheapest per provider shown). All A16 offers, price history and alerts
V100
| Provider | Region | GPUs | Per GPU / hr | Instance / hr | Deploy |
|---|---|---|---|---|---|
| Ori | beauharnois | 1 | $0.83 | — | Deploy |
| Paperspace | ca1 | 1 | $2.30 | — | Deploy |
2 providers in stock, 72 offers (cheapest per provider shown). All V100 offers, price history and alerts
Notify me when A16 drops below a price
One email when the cheapest in-stock on-demand price per GPU-hour falls below your threshold. Cheapest right now: $0.06/GPU-hr.
QuantaCloud
Comparing providers? We broker across all of them.
Stop tab-switching between pricing pages. Tell us what you need, 16+ GPUs reserved or cluster capacity, and we return one quote at partner rates within 24 hours.
When to Choose the A16
The A16 suits inference workloads that utilize its INT8 dense performance of 35.9 TOPS. Its 250W TDP supports deployments where power limits matter more than peak FP16 dense throughput. PCIe 4.0 interconnect fits systems without NVLink requirements.
When to Choose the V100
The V100 suits training workloads that utilize its FP16 dense performance of 125 TFLOPS. Its 900 GB/s memory bandwidth supports larger batch sizes than the 200 GB/s on the A16. NVLink interconnect fits multi GPU scaling needs.
Use Cases
The V100 delivers 125 TFLOPS FP16 dense compared to 17.9 TFLOPS on the A16.
The V100 delivers 125 TFLOPS FP16 dense compared to 17.9 TFLOPS on the A16.
The V100 delivers 15.7 TFLOPS FP32 compared to 4.5 TFLOPS on the A16.
The V100 delivers 125 TFLOPS FP16 dense compared to 17.9 TFLOPS on the A16.
The V100 delivers 15.7 TFLOPS FP32 compared to 4.5 TFLOPS on the A16.
Frequently Asked Questions
What is the FP16 dense performance of the A16?▾
The A16 lists FP16 dense performance at 17.9 TFLOPS. It also lists FP16 with sparsity at 35.9 TFLOPS.
What is the FP16 dense performance of the V100?▾
The V100 lists FP16 dense performance at 125 TFLOPS. No sparsity figure appears in the specifications.
How does memory bandwidth compare between the A16 and V100?▾
The A16 provides 200 GB/s bandwidth. The V100 provides 900 GB/s bandwidth.
What TDP values apply to the A16 and V100?▾
The A16 has a TDP of 250W. The V100 has a TDP of 300W.
Which interconnects does each GPU support?▾
The A16 supports PCIe 4.0. The V100 supports NVLink and PCIe 3.0.
What FP32 performance does each GPU deliver?▾
The A16 delivers 4.5 TFLOPS FP32. The V100 delivers 15.7 TFLOPS FP32.
Which is cheaper to rent, the A16 or the V100?▾
Cloud rental prices for both the A16 and V100 vary by provider, configuration, and availability. This page shows live pricing from 25+ providers updated every 60 seconds. Scroll to the Live Cloud Pricing section to compare current rates.
How much VRAM does the A16 have compared to the V100?▾
The A16 has 16 GB of GDDR6 memory. The V100 has 16 to 32 GB of HBM2 memory.
Can I find A16 and V100 GPUs available to rent right now?▾
Yes. This page shows real-time availability across 25+ cloud GPU providers. The Live Cloud Pricing section displays only in-stock offers with current pricing.
What is the main difference between the A16 and the V100?▾
The A16 uses the Ampere architecture (2021) while the V100 uses Volta (2017). The V100 delivers 7.0x the dense FP16 throughput (125 vs 17.9 TFLOPS, both without sparsity) and 4.5x the memory bandwidth of the A16.
Rent these GPUs
Each GPU page lists every current on-demand offer by provider, per GPU-hour, updated every minute.
Related comparisons
How this page is made
- Specifications come from the NVIDIA, AMD and Intel datasheets for the A16 and the V100. Dense and sparse throughput are listed separately.
- Prices and availability are live: every offer shown is an in-stock, on-demand listing from a provider we track, rechecked every minute.
- The written comparison (overview, performance notes, when to choose each card, verdict, use cases and FAQ) was drafted with an AI model from the specification table above and passed an automated check that rejects any figure not in that table. It was last generated on . It contains no prices; those are always read live.
- Read how we collect the data, or report an error on this page.
Next steps
- Rent A16Every current offer by provider, daily price history and a price alert.
- Rent V100Every current offer by provider, daily price history and a price alert.
- GPU price indexHow on-demand prices for the major GPUs have moved, updated daily.
- All GPUsEvery GPU we track, with current lows and provider counts.