L40 vs RTX 6000 Ada

Ada LovelacevsAda LovelaceUpdated 7 days ago

The RTX 6000 Ada wins for the most common use case of LLM training and inference. Its 364 TFLOPS dense FP16 rating doubles the corresponding 181.05 TFLOPS rating on the L40 while memory capacity and TDP remain identical.

L40 from $0.82/GPU/hrRTX 6000 Ada from $0.78/GPU/hr

Right now, from live stock

  • Cheapest right now: RTX 6000 Ada at $0.78/hr on QuantaCloud

    Deploy
  • Most providers in stock: RTX 6000 Ada (5)

    See all offers
  • Best $ per TFLOPS FP16 dense right now: RTX 6000 Ada ($0.0021 per TFLOPS-hour at 364 TFLOPS)

    Deploy

Specifications Compared

SpecL40RTX 6000 Ada
TDP300W300W
VRAM48 GB48 GB
CUDA Cores18,17618,176
FP8 (dense)362 TFLOPS728.5 TFLOPS
Memory TypeGDDR6GDDR6
ArchitectureAda LovelaceAda Lovelace
FP16 (dense)181.05 TFLOPS364 TFLOPS
Form FactorsPCIePCIe
INT8 (dense)362 TOPS728.5 TOPS
InterconnectPCIe 4.0PCIe 4.0
Tensor Cores568568
FP32 Performance90.5 TFLOPS91.1 TFLOPS
Memory Bandwidth864 GB/s960 GB/s
FP8 (with sparsity)724 TFLOPS1,457 TFLOPS
FP16 (with sparsity)362.1 TFLOPS728 TFLOPS
INT8 (with sparsity)724 TOPS1,457 TOPS

Performance Analysis

Dense FP16 throughput on the RTX 6000 Ada measures 364 TFLOPS against 181.05 TFLOPS on the L40. Training and inference steps that depend on dense FP16 arithmetic therefore finish in approximately half the time on the RTX 6000 Ada. The FP32 ratings remain nearly identical at 90.5 TFLOPS for the L40 and 91.1 TFLOPS for the RTX 6000 Ada. Memory bandwidth of 960 GB/s on the RTX 6000 Ada exceeds the 864 GB/s figure on the L40. The added bandwidth permits modestly larger batch sizes when data movement constrains workload speed. Sparse FP16 performance follows the identical ratio with 728 TFLOPS on the RTX 6000 Ada versus 362.1 TFLOPS on the L40.

Current On-Demand Offers

Cheapest secure on-demand offer per provider that is reported in stock right now, per GPU per hour; multi-GPU instances show the whole-instance price alongside. Deploy opens the DeployGPU console for providers on the platform, otherwise the provider's own site. Stock is rechecked every minute.

L40

ProviderRegionGPUsPer GPU / hrInstance / hrDeploy
RunPodglobal1$0.82Deploy
Massed Computeus-central-31$0.86Deploy
QuantaCloudus-midwest-24$0.94$3.75Deploy
HyperstackCANADA-12$1.00$2.00Deploy

4 providers in stock, 10 offers (cheapest per provider shown). All L40 offers, price history and alerts

RTX 6000 Ada

ProviderRegionGPUsPer GPU / hrInstance / hrDeploy
QuantaCloudus-midwest-14$0.78$3.11Deploy
Massed Computeus-central-31$0.79Deploy
RunPodglobal1$0.84Deploy
LeaderGPUThe Netherlands8$1.20$9.60Deploy
DigitalOceanToronto 11$1.57Deploy

5 providers in stock, 15 offers (cheapest per provider shown). All RTX 6000 Ada offers, price history and alerts

Which GPU to watchWatch the price of

Notify me when L40 drops below a price

One email when the cheapest in-stock on-demand price per GPU-hour falls below your threshold. Cheapest right now: $0.82/GPU-hr.

QuantaCloud

Comparing providers? We broker across all of them.

Stop tab-switching between pricing pages. Tell us what you need, 16+ GPUs reserved or cluster capacity, and we return one quote at partner rates within 24 hours.

No waitlist24hr quote turnaroundInfiniBand fabric

Compare real-time pricing across 25+ providers

When to Choose the L40

The L40 suits scientific computing workloads that rely primarily on the 90.5 TFLOPS FP32 rating. In such cases the nearly identical FP32 number relative to the RTX 6000 Ada reduces the incentive to select the higher tensor throughput card.

When to Choose the RTX 6000 Ada

The RTX 6000 Ada suits LLM training and inference that exploit the 364 TFLOPS dense FP16 rating. The doubled dense FP16 throughput relative to the L40 shortens iteration times when tensor operations dominate execution.

Use Cases

LLM Training
RTX 6000 Ada

The RTX 6000 Ada supplies 364 TFLOPS dense FP16 compared with 181.05 TFLOPS on the L40.

LLM Inference
RTX 6000 Ada

The RTX 6000 Ada supplies 364 TFLOPS dense FP16 compared with 181.05 TFLOPS on the L40.

Fine-tuning
RTX 6000 Ada

The RTX 6000 Ada supplies 364 TFLOPS dense FP16 compared with 181.05 TFLOPS on the L40.

Stable Diffusion
RTX 6000 Ada

The RTX 6000 Ada supplies 364 TFLOPS dense FP16 compared with 181.05 TFLOPS on the L40.

Scientific Computing
Either

FP32 performance measures 90.5 TFLOPS on the L40 and 91.1 TFLOPS on the RTX 6000 Ada.

Frequently Asked Questions

What are the dense FP16 ratings for each card?

The L40 reaches 181.05 TFLOPS dense FP16 while the RTX 6000 Ada reaches 364 TFLOPS dense FP16.

How does memory bandwidth differ between the two GPUs?

The L40 provides 864 GB/s memory bandwidth while the RTX 6000 Ada provides 960 GB/s memory bandwidth.

What FP32 performance do the cards deliver?

The L40 delivers 90.5 TFLOPS FP32 while the RTX 6000 Ada delivers 91.1 TFLOPS FP32.

Do the GPUs share the same TDP and interconnect?

Both the L40 and the RTX 6000 Ada operate at 300 W TDP with PCIe 4.0 interconnect.

How do the sparse FP16 figures compare?

The L40 reaches 362.1 TFLOPS with sparsity in FP16 while the RTX 6000 Ada reaches 728 TFLOPS with sparsity in FP16.

Which is cheaper to rent, the L40 or the RTX 6000 Ada?

Cloud rental prices for both the L40 and RTX 6000 Ada vary by provider, configuration, and availability. This page shows live pricing from 25+ providers updated every 60 seconds. Scroll to the Live Cloud Pricing section to compare current rates.

How much VRAM does the L40 have compared to the RTX 6000 Ada?

The L40 has 48 GB of GDDR6 memory. The RTX 6000 Ada has 48 GB of GDDR6 memory.

Can I find L40 and RTX 6000 Ada GPUs available to rent right now?

Yes. This page shows real-time availability across 25+ cloud GPU providers. The Live Cloud Pricing section displays only in-stock offers with current pricing.

What is the main difference between the L40 and the RTX 6000 Ada?

The L40 uses the Ada Lovelace architecture (2023) while the RTX 6000 Ada uses Ada Lovelace (2022). The RTX 6000 Ada delivers 2.0x the dense FP16 throughput (364 vs 181.05 TFLOPS, both without sparsity) and 1.1x the memory bandwidth of the L40.

Rent these GPUs

Each GPU page lists every current on-demand offer by provider, per GPU-hour, updated every minute.

Related comparisons

How this page is made

  • Specifications come from the NVIDIA, AMD and Intel datasheets for the L40 and the RTX 6000 Ada. Dense and sparse throughput are listed separately.
  • Prices and availability are live: every offer shown is an in-stock, on-demand listing from a provider we track, rechecked every minute.
  • The written comparison (overview, performance notes, when to choose each card, verdict, use cases and FAQ) was drafted with an AI model from the specification table above and passed an automated check that rejects any figure not in that table. It was last generated on . It contains no prices; those are always read live.
  • Read how we collect the data, or report an error on this page.

Next steps