Specifications Compared
| Spec | L4 | RTX 4090 |
|---|---|---|
| TDP | 72W | 450W |
| VRAM | 24 GB | 24 GB |
| CUDA Cores | 7,424 | 16,384 |
| FP8 (dense) | 242 TFLOPS | 330.3 TFLOPS |
| Memory Type | GDDR6 | GDDR6X |
| Architecture | Ada Lovelace | Ada Lovelace |
| FP16 (dense) | 121 TFLOPS | 165.2 TFLOPS |
| Form Factors | PCIe | PCIe |
| INT8 (dense) | 242 TOPS | 660.6 TOPS |
| Interconnect | PCIe 4.0 | PCIe 4.0 |
| Tensor Cores | 232 | 512 |
| FP32 Performance | 30.3 TFLOPS | 82.6 TFLOPS |
| FP64 Performance | 0.5 TFLOPS | 1.3 TFLOPS |
| Memory Bandwidth | 300 GB/s | 1,008 GB/s |
| FP8 (with sparsity) | 485 TFLOPS | 660.6 TFLOPS |
| FP16 (with sparsity) | 242 TFLOPS | 330.4 TFLOPS |
| INT8 (with sparsity) | 485 TOPS | 1,321.2 TOPS |
Performance Analysis
Dense FP16 performance measures 121 TFLOPS on the L4 and 165.2 TFLOPS on the RTX 4090. The same dense metric for FP8 shows 242 TFLOPS versus 330.3 TFLOPS. These gaps indicate that the RTX 4090 sustains higher throughput during both training and inference workloads that rely on dense tensor operations. The FP16 to FP32 ratio equals four times on the L4 and two times on the RTX 4090. Memory bandwidth of 300 GB/s on the L4 versus 1008 GB/s on the RTX 4090 directly limits sustainable batch sizes during large model execution. Sparse FP16 figures of 242 TFLOPS and 330.4 TFLOPS preserve the same relative ordering between the two cards.
Current On-Demand Offers
Cheapest secure on-demand offer per provider that is reported in stock right now, per GPU per hour; multi-GPU instances show the whole-instance price alongside. Deploy opens the DeployGPU console for providers on the platform, otherwise the provider's own site. Stock is rechecked every minute.
L4
| Provider | Region | GPUs | Per GPU / hr | Instance / hr | Deploy |
|---|---|---|---|---|---|
| Scaleway | Warsaw, Poland (WAW2) | 1 | $0.90 | — | Deploy |
| Ori | lille-4 | 1 | $0.93 | — | Deploy |
2 providers in stock, 2 offers. All L4 offers, price history and alerts
Notify me when L4 drops below a price
One email when the cheapest in-stock on-demand price per GPU-hour falls below your threshold. Cheapest right now: $0.90/GPU-hr.
QuantaCloud
Comparing providers? We broker across all of them.
Stop tab-switching between pricing pages. Tell us what you need, 16+ GPUs reserved or cluster capacity, and we return one quote at partner rates within 24 hours.
When to Choose the L4
The L4 suits deployments where total power consumption must remain near 72W. Edge inference nodes and dense server racks benefit from the 121 TFLOPS dense FP16 rating paired with the low TDP. Workloads that fit within 24 GB yet do not require maximum bandwidth also favor the L4.
When to Choose the RTX 4090
The RTX 4090 fits training sessions that exploit its 82.6 TFLOPS FP32 and 165.2 TFLOPS dense FP16. Larger batch sizes become feasible with the 1008 GB/s bandwidth. Users running dense FP8 operations at 330.3 TFLOPS likewise select the RTX 4090.
Use Cases
The RTX 4090 supplies 82.6 TFLOPS FP32 and 165.2 TFLOPS dense FP16 against the L4 values of 30.3 TFLOPS and 121 TFLOPS.
Higher dense FP8 throughput of 330.3 TFLOPS and 1008 GB/s bandwidth support larger concurrent request volumes than the L4.
The 165.2 TFLOPS dense FP16 figure and 1008 GB/s bandwidth enable faster iteration than the L4 configuration.
The RTX 4090 delivers 330.3 TFLOPS dense FP8 and 1008 GB/s bandwidth for accelerated image generation steps.
FP32 performance of 82.6 TFLOPS exceeds the L4 rating of 30.3 TFLOPS for double-precision adjacent workloads.
Frequently Asked Questions
How does memory bandwidth differ between the L4 and RTX 4090?▾
The L4 provides 300 GB/s while the RTX 4090 reaches 1008 GB/s. This difference influences maximum batch size during model execution.
What FP16 dense performance does each card offer?▾
The L4 lists 121 TFLOPS dense FP16. The RTX 4090 lists 165.2 TFLOPS dense FP16.
Which card consumes less power?▾
The L4 operates at a 72W TDP. The RTX 4090 operates at a 450W TDP.
How do the cards compare on sparse FP16?▾
The L4 reaches 242 TFLOPS with sparsity. The RTX 4090 reaches 330.4 TFLOPS with sparsity.
Are the interconnects identical?▾
Both cards employ PCIe 4.0 interconnects in PCIe form factor.
Which is cheaper to rent, the L4 or the RTX 4090?▾
Cloud rental prices for both the L4 and RTX 4090 vary by provider, configuration, and availability. This page shows live pricing from 25+ providers updated every 60 seconds. Scroll to the Live Cloud Pricing section to compare current rates.
How much VRAM does the L4 have compared to the RTX 4090?▾
The L4 has 24 GB of GDDR6 memory. The RTX 4090 has 24 GB of GDDR6X memory.
Can I find L4 and RTX 4090 GPUs available to rent right now?▾
Yes. This page shows real-time availability across 25+ cloud GPU providers. The Live Cloud Pricing section displays only in-stock offers with current pricing.
What is the main difference between the L4 and the RTX 4090?▾
The L4 uses the Ada Lovelace architecture (2023) while the RTX 4090 uses Ada Lovelace (2022). The RTX 4090 delivers 1.4x the dense FP16 throughput (165.2 vs 121 TFLOPS, both without sparsity) and 3.4x the memory bandwidth of the L4.
Rent these GPUs
Each GPU page lists every current on-demand offer by provider, per GPU-hour, updated every minute.
Related comparisons
How this page is made
- Specifications come from the NVIDIA, AMD and Intel datasheets for the L4 and the RTX 4090. Dense and sparse throughput are listed separately.
- Prices and availability are live: every offer shown is an in-stock, on-demand listing from a provider we track, rechecked every minute.
- The written comparison (overview, performance notes, when to choose each card, verdict, use cases and FAQ) was drafted with an AI model from the specification table above and passed an automated check that rejects any figure not in that table. It was last generated on . It contains no prices; those are always read live.
- Read how we collect the data, or report an error on this page.
Next steps
- Rent L4Every current offer by provider, daily price history and a price alert.
- Rent RTX 4090Every current offer by provider, daily price history and a price alert.
- GPU price indexHow on-demand prices for the major GPUs have moved, updated daily.
- All GPUsEvery GPU we track, with current lows and provider counts.