Specifications Compared
| Spec | L40S | RTX A6000 |
|---|---|---|
| TDP | 350W | 300W |
| VRAM | 48 GB | 48 GB |
| CUDA Cores | 18,176 | 10,752 |
| FP8 (dense) | 733 TFLOPS | Not published |
| Memory Type | GDDR6 | GDDR6 |
| Architecture | Ada Lovelace | Ampere |
| FP16 (dense) | 362.05 TFLOPS | 154.8 TFLOPS |
| Form Factors | PCIe | PCIe |
| INT8 (dense) | 733 TOPS | 309.7 TOPS |
| Interconnect | PCIe 4.0 | NVLink, PCIe 4.0 |
| Tensor Cores | 568 | 336 |
| FP32 Performance | 91.6 TFLOPS | 38.7 TFLOPS |
| FP64 Performance | 1.4 TFLOPS | Not published |
| Memory Bandwidth | 864 GB/s | 768 GB/s |
| FP8 (with sparsity) | 1,466 TFLOPS | Not published |
| FP16 (with sparsity) | 733 TFLOPS | 309.6 TFLOPS |
| INT8 (with sparsity) | 1,466 TOPS | 619.4 TOPS |
Performance Analysis
Dense FP16 performance reaches 362.05 TFLOPS on the L40S versus 154.8 TFLOPS on the RTX A6000. This gap means training and inference workloads that use dense FP16 operations complete more operations per second on the L40S. Sparse FP16 performance reaches 733 TFLOPS on the L40S versus 309.6 TFLOPS on the RTX A6000 so workloads that activate sparsity see a similar proportional advantage. FP32 performance stands at 91.6 TFLOPS for the L40S and 38.7 TFLOPS for the RTX A6000. Memory bandwidth of 864 GB/s on the L40S exceeds 768 GB/s on the RTX A6000 and therefore supports modestly larger batch sizes in memory bound phases.
Current On-Demand Offers
Cheapest secure on-demand offer per provider that is reported in stock right now, per GPU per hour; multi-GPU instances show the whole-instance price alongside. Deploy opens the DeployGPU console for providers on the platform, otherwise the provider's own site. Stock is rechecked every minute.
L40S
| Provider | Region | GPUs | Per GPU / hr | Instance / hr | Deploy |
|---|---|---|---|---|---|
| Vast.ai | Slovenia, SI | 1 | $0.80 | — | Deploy |
| Massed Compute | us-central-3 | 1 | $0.97 | — | Deploy |
| QuantaCloud | us-midwest-1 | 4 | $1.09 | $4.36 | Deploy |
| RunPod | global | 1 | $1.09 | — | Deploy |
| Lyceum | Europe | 2 | $1.19 | $2.38 | Deploy |
6 providers in stock, 14 offers (cheapest per provider shown). All L40S offers, price history and alerts
RTX A6000
| Provider | Region | GPUs | Per GPU / hr | Instance / hr | Deploy |
|---|---|---|---|---|---|
| LeaderGPU | The Netherlands | 8 | $0.44 | $3.54 | Deploy |
| QuantaCloud | us-midwest-2 | 2 | $0.48 | $0.96 | Deploy |
| Hyperstack | CANADA-1 | 4 | $0.50 | $2.00 | Deploy |
| RunPod | global | 1 | $0.53 | — | Deploy |
| Massed Compute | us-central-2 | 4 | $0.55 | $2.20 | Deploy |
7 providers in stock, 58 offers (cheapest per provider shown). All RTX A6000 offers, price history and alerts
Notify me when L40S drops below a price
One email when the cheapest in-stock on-demand price per GPU-hour falls below your threshold. Cheapest right now: $0.80/GPU-hr.
QuantaCloud
Comparing providers? We broker across all of them.
Stop tab-switching between pricing pages. Tell us what you need, 16+ GPUs reserved or cluster capacity, and we return one quote at partner rates within 24 hours.
When to Choose the L40S
The L40S serves workloads that require 362.05 TFLOPS dense FP16 or 733 TFLOPS sparse FP16. Its 864 GB/s memory bandwidth further aids data movement in large batch inference. The 91.6 TFLOPS FP32 rating also exceeds the alternative by more than double.
When to Choose the RTX A6000
The RTX A6000 serves workloads that require the NVLink interconnect option. Its 300W TDP rating is lower than 350W and therefore reduces power draw in constrained environments. The 48 GB GDDR6 capacity remains identical to the L40S.
Use Cases
The L40S supplies 362.05 TFLOPS dense FP16 which exceeds the 154.8 TFLOPS dense FP16 of the RTX A6000.
The L40S supplies 362.05 TFLOPS dense FP16 which exceeds the 154.8 TFLOPS dense FP16 of the RTX A6000.
The L40S supplies 362.05 TFLOPS dense FP16 which exceeds the 154.8 TFLOPS dense FP16 of the RTX A6000.
The L40S supplies 362.05 TFLOPS dense FP16 which exceeds the 154.8 TFLOPS dense FP16 of the RTX A6000.
Both GPUs provide 48 GB GDDR6 yet the L40S offers higher FP32 performance at 91.6 TFLOPS versus 38.7 TFLOPS.
Frequently Asked Questions
How does dense FP16 performance compare between the L40S and RTX A6000?▾
The L40S lists 362.05 TFLOPS dense FP16 while the RTX A6000 lists 154.8 TFLOPS dense FP16.
What is the FP32 performance of each GPU?▾
The L40S lists 91.6 TFLOPS FP32 and the RTX A6000 lists 38.7 TFLOPS FP32.
Which GPU has higher memory bandwidth?▾
The L40S provides 864 GB/s memory bandwidth while the RTX A6000 provides 768 GB/s memory bandwidth.
What interconnect options exist on the RTX A6000?▾
The RTX A6000 supports NVLink in addition to PCIe 4.0 while the L40S supports only PCIe 4.0.
How do the TDP ratings differ?▾
The L40S has a TDP of 350W and the RTX A6000 has a TDP of 300W.
Which is cheaper to rent, the L40S or the RTX A6000?▾
Cloud rental prices for both the L40S and RTX A6000 vary by provider, configuration, and availability. This page shows live pricing from 25+ providers updated every 60 seconds. Scroll to the Live Cloud Pricing section to compare current rates.
How much VRAM does the L40S have compared to the RTX A6000?▾
The L40S has 48 GB of GDDR6 memory. The RTX A6000 has 48 GB of GDDR6 memory.
Can I find L40S and RTX A6000 GPUs available to rent right now?▾
Yes. This page shows real-time availability across 25+ cloud GPU providers. The Live Cloud Pricing section displays only in-stock offers with current pricing.
What is the main difference between the L40S and the RTX A6000?▾
The L40S uses the Ada Lovelace architecture (2023) while the RTX A6000 uses Ampere (2020). The L40S delivers 2.3x the dense FP16 throughput (362.05 vs 154.8 TFLOPS, both without sparsity) and 1.1x the memory bandwidth of the RTX A6000.
Rent these GPUs
Each GPU page lists every current on-demand offer by provider, per GPU-hour, updated every minute.
Related comparisons
How this page is made
- Specifications come from the NVIDIA, AMD and Intel datasheets for the L40S and the RTX A6000. Dense and sparse throughput are listed separately.
- Prices and availability are live: every offer shown is an in-stock, on-demand listing from a provider we track, rechecked every minute.
- The written comparison (overview, performance notes, when to choose each card, verdict, use cases and FAQ) was drafted with an AI model from the specification table above and passed an automated check that rejects any figure not in that table. It was last generated on . It contains no prices; those are always read live.
- Read how we collect the data, or report an error on this page.
Next steps
- Rent L40SEvery current offer by provider, daily price history and a price alert.
- Rent RTX A6000Every current offer by provider, daily price history and a price alert.
- GPU price indexHow on-demand prices for the major GPUs have moved, updated daily.
- All GPUsEvery GPU we track, with current lows and provider counts.