Specifications Compared
| Spec | L40 | RTX 6000 Ada |
|---|---|---|
| TDP | 300W | 300W |
| VRAM | 48 GB | 48 GB |
| CUDA Cores | 18,176 | 18,176 |
| FP8 (dense) | 362 TFLOPS | 728.5 TFLOPS |
| Memory Type | GDDR6 | GDDR6 |
| Architecture | Ada Lovelace | Ada Lovelace |
| FP16 (dense) | 181.05 TFLOPS | 364 TFLOPS |
| Form Factors | PCIe | PCIe |
| INT8 (dense) | 362 TOPS | 728.5 TOPS |
| Interconnect | PCIe 4.0 | PCIe 4.0 |
| Tensor Cores | 568 | 568 |
| FP32 Performance | 90.5 TFLOPS | 91.1 TFLOPS |
| Memory Bandwidth | 864 GB/s | 960 GB/s |
| FP8 (with sparsity) | 724 TFLOPS | 1,457 TFLOPS |
| FP16 (with sparsity) | 362.1 TFLOPS | 728 TFLOPS |
| INT8 (with sparsity) | 724 TOPS | 1,457 TOPS |
Performance Analysis
Dense FP16 throughput on the RTX 6000 Ada measures 364 TFLOPS against 181.05 TFLOPS on the L40. Training and inference steps that depend on dense FP16 arithmetic therefore finish in approximately half the time on the RTX 6000 Ada. The FP32 ratings remain nearly identical at 90.5 TFLOPS for the L40 and 91.1 TFLOPS for the RTX 6000 Ada. Memory bandwidth of 960 GB/s on the RTX 6000 Ada exceeds the 864 GB/s figure on the L40. The added bandwidth permits modestly larger batch sizes when data movement constrains workload speed. Sparse FP16 performance follows the identical ratio with 728 TFLOPS on the RTX 6000 Ada versus 362.1 TFLOPS on the L40.
Current On-Demand Offers
Cheapest secure on-demand offer per provider that is reported in stock right now, per GPU per hour; multi-GPU instances show the whole-instance price alongside. Deploy opens the DeployGPU console for providers on the platform, otherwise the provider's own site. Stock is rechecked every minute.
L40
| Provider | Region | GPUs | Per GPU / hr | Instance / hr | Deploy |
|---|---|---|---|---|---|
| RunPod | global | 1 | $0.82 | — | Deploy |
| Massed Compute | us-central-3 | 1 | $0.86 | — | Deploy |
| QuantaCloud | us-midwest-2 | 4 | $0.94 | $3.75 | Deploy |
| Hyperstack | CANADA-1 | 2 | $1.00 | $2.00 | Deploy |
4 providers in stock, 10 offers (cheapest per provider shown). All L40 offers, price history and alerts
RTX 6000 Ada
| Provider | Region | GPUs | Per GPU / hr | Instance / hr | Deploy |
|---|---|---|---|---|---|
| QuantaCloud | us-midwest-1 | 4 | $0.78 | $3.11 | Deploy |
| Massed Compute | us-central-3 | 1 | $0.79 | — | Deploy |
| RunPod | global | 1 | $0.84 | — | Deploy |
| LeaderGPU | The Netherlands | 8 | $1.20 | $9.60 | Deploy |
| DigitalOcean | Toronto 1 | 1 | $1.57 | — | Deploy |
5 providers in stock, 15 offers (cheapest per provider shown). All RTX 6000 Ada offers, price history and alerts
Notify me when L40 drops below a price
One email when the cheapest in-stock on-demand price per GPU-hour falls below your threshold. Cheapest right now: $0.82/GPU-hr.
QuantaCloud
Comparing providers? We broker across all of them.
Stop tab-switching between pricing pages. Tell us what you need, 16+ GPUs reserved or cluster capacity, and we return one quote at partner rates within 24 hours.
When to Choose the L40
The L40 suits scientific computing workloads that rely primarily on the 90.5 TFLOPS FP32 rating. In such cases the nearly identical FP32 number relative to the RTX 6000 Ada reduces the incentive to select the higher tensor throughput card.
When to Choose the RTX 6000 Ada
The RTX 6000 Ada suits LLM training and inference that exploit the 364 TFLOPS dense FP16 rating. The doubled dense FP16 throughput relative to the L40 shortens iteration times when tensor operations dominate execution.
Use Cases
The RTX 6000 Ada supplies 364 TFLOPS dense FP16 compared with 181.05 TFLOPS on the L40.
The RTX 6000 Ada supplies 364 TFLOPS dense FP16 compared with 181.05 TFLOPS on the L40.
The RTX 6000 Ada supplies 364 TFLOPS dense FP16 compared with 181.05 TFLOPS on the L40.
The RTX 6000 Ada supplies 364 TFLOPS dense FP16 compared with 181.05 TFLOPS on the L40.
FP32 performance measures 90.5 TFLOPS on the L40 and 91.1 TFLOPS on the RTX 6000 Ada.
Frequently Asked Questions
What are the dense FP16 ratings for each card?▾
The L40 reaches 181.05 TFLOPS dense FP16 while the RTX 6000 Ada reaches 364 TFLOPS dense FP16.
How does memory bandwidth differ between the two GPUs?▾
The L40 provides 864 GB/s memory bandwidth while the RTX 6000 Ada provides 960 GB/s memory bandwidth.
What FP32 performance do the cards deliver?▾
The L40 delivers 90.5 TFLOPS FP32 while the RTX 6000 Ada delivers 91.1 TFLOPS FP32.
Do the GPUs share the same TDP and interconnect?▾
Both the L40 and the RTX 6000 Ada operate at 300 W TDP with PCIe 4.0 interconnect.
How do the sparse FP16 figures compare?▾
The L40 reaches 362.1 TFLOPS with sparsity in FP16 while the RTX 6000 Ada reaches 728 TFLOPS with sparsity in FP16.
Which is cheaper to rent, the L40 or the RTX 6000 Ada?▾
Cloud rental prices for both the L40 and RTX 6000 Ada vary by provider, configuration, and availability. This page shows live pricing from 25+ providers updated every 60 seconds. Scroll to the Live Cloud Pricing section to compare current rates.
How much VRAM does the L40 have compared to the RTX 6000 Ada?▾
The L40 has 48 GB of GDDR6 memory. The RTX 6000 Ada has 48 GB of GDDR6 memory.
Can I find L40 and RTX 6000 Ada GPUs available to rent right now?▾
Yes. This page shows real-time availability across 25+ cloud GPU providers. The Live Cloud Pricing section displays only in-stock offers with current pricing.
What is the main difference between the L40 and the RTX 6000 Ada?▾
The L40 uses the Ada Lovelace architecture (2023) while the RTX 6000 Ada uses Ada Lovelace (2022). The RTX 6000 Ada delivers 2.0x the dense FP16 throughput (364 vs 181.05 TFLOPS, both without sparsity) and 1.1x the memory bandwidth of the L40.
Rent these GPUs
Each GPU page lists every current on-demand offer by provider, per GPU-hour, updated every minute.
Related comparisons
How this page is made
- Specifications come from the NVIDIA, AMD and Intel datasheets for the L40 and the RTX 6000 Ada. Dense and sparse throughput are listed separately.
- Prices and availability are live: every offer shown is an in-stock, on-demand listing from a provider we track, rechecked every minute.
- The written comparison (overview, performance notes, when to choose each card, verdict, use cases and FAQ) was drafted with an AI model from the specification table above and passed an automated check that rejects any figure not in that table. It was last generated on . It contains no prices; those are always read live.
- Read how we collect the data, or report an error on this page.
Next steps
- Rent L40Every current offer by provider, daily price history and a price alert.
- Rent RTX 6000 ADAEvery current offer by provider, daily price history and a price alert.
- GPU price indexHow on-demand prices for the major GPUs have moved, updated daily.
- All GPUsEvery GPU we track, with current lows and provider counts.