NVIDIAAda Lovelace Architecture

Rent NVIDIA L40S

The NVIDIA L40S GPU is a data center GPU designed for demanding visualization, compute, and AI workloads. It offers a balance of performance and features, making it suitable for a wide range of professional applications.
Current on-demand floor$0.97

per GPU-hour

📊 Pricing at a Glance

Cheapest In Stock
Massed Compute
$0.97/GPU/hr
Most Expensive In Stock
Ori
$1.55/GPU/hr
Median In-Stock Price
$1.09/GPU/hr
Offers Reported In Stock
12
Providers With Stock
5
Latest Stock Observation

NVIDIA L40S on-demand pricing, in stock right now, ranges from $0.97/GPU/hr to $1.55/GPU/hr across 12 offers from 5 providers with live stock tracking (updated October 2026). A further 107 tracked listings are shown below as listed / last-seen pricing.

Looking for a specific provider? See RunPod NVIDIA L40S, Lyceum NVIDIA L40S, or TensorDock NVIDIA L40S.

Available Offers

The 5 cheapest on-demand offers reported in stock right now, from 5 providers with live stock tracking.

12 offers available
NVIDIA L40S2x
48GB VRAM
24 vCPU
144GB RAM
1250GB Storage
$0.97/GPU/hr
$1.94/hr total (2×)
NVIDIA L40S
48GB VRAM
12 vCPU
72GB RAM
625GB Storage
$0.97/GPU/hr
NVIDIA L40S4x
48GB VRAM
46 vCPU
288GB RAM
2500GB Storage
$0.97/GPU/hr
$3.88/hr total (4×)
NVIDIA L40S
48GB VRAM
12 vCPU
72GB RAM
625GB Storage
$0.97/GPU/hr
NVIDIA L40S
48GB VRAM
16 vCPU
94GB RAM
$1.09/GPU/hr

Other Tracked Listings (Last-Seen Pricing)

107 further listings GPUPerHour tracks for this GPU. These catalog rows do not qualify as current secure on-demand offers; stock observations may be missing or old. They are not confirmed in stock; check availability with the provider.

ProviderListed per GPU / hrPricing typeAvailability
TensorDock$0.47on-demandCheck provider
Crusoe$0.50on-demandCheck provider
TensorDock$0.55on-demandCheck provider
TensorDock$0.55on-demandCheck provider
Latitude.sh$0.74on-demandCheck provider
Nebius$0.74on-demandCheck provider
Latitude.sh$0.74on-demandCheck provider
Vast.ai$0.80on-demandCheck provider
Vast.ai$0.80on-demandCheck provider
Vast.ai$0.80on-demandCheck provider
Browse all tracked listings

L40S Availability Now

5 providers report L40S in stock on-demand right now (12 offers across 7 regions). Last stock observation: .

Cheapest in-stock L40S offer per region
RegionCheapest $/GPU-hrProviderOffers in stock
us-central-2$0.97Massed Compute3
us-central-3$0.97Massed Compute1
global$1.09RunPod1
us-midwest-2$1.09QuantaCloud1
us-midwest-1$1.09QuantaCloud3

The 5 cheapest of 7 regions with stock. Regions are as the providers name them. Prices on this page were observed on (UTC) and are quoted for that day; they are rechecked every minute and can change before booking.

L40S Price History (daily)

Over the last 14 days the cheapest on-demand L40S ranged from $0.80 to $0.97 per GPU-hour across up to 8 providers with stock.

L40S daily cheapest and median on-demand price per GPU-hour, last 14 days
Date (UTC)Cheapest $/GPU-hrMedian $/GPU-hrOffers in stockProviders with stock
$0.80$1.09136
$0.97$1.09175
$0.97$1.09185
$0.80$1.1943
$0.97$1.09186
$0.97$1.09186
$0.92$1.09227
$0.80$1.09228
$0.92$1.09196
$0.80$1.09177
$0.97$1.09165
$0.97$1.09175
$0.97$1.09216
$0.97$1.09206

One row per captured day (00:05 UTC). Cheapest is the lowest secure on-demand price per GPU-hour reported in stock by a provider with live stock tracking; median is the median of those offers. Days without a capture are not shown. Method and full series: Cloud GPU Price Index.

L40S Price by Provider

Cheapest on-demand offer per provider that is reported in stock right now. Prices are per GPU per hour; multi-GPU instances show the whole-instance price alongside. Deploy opens the DeployGPU console for providers on the platform, otherwise the provider's own site.

ProviderRegionGPUsPer GPU / hrInstance / hrDeploy
Massed Computeus-central-22$0.97$1.94Deploy this GPU
RunPodglobal1$1.09$1.09Deploy this GPU
QuantaCloudus-midwest-21$1.09$1.09Deploy this GPU
LyceumEurope1$1.19$1.19Deploy this GPU
Orilille-41$1.55$1.55Deploy this GPU

Notify me when L40S drops below a price

One email when the cheapest in-stock on-demand price per GPU-hour falls below your threshold. Cheapest right now: $0.97/GPU-hr.

L40S Catalog Listing History (30 days)

Catalog listings, not live prices.

These averages are of every tracked catalog listing, including spot, reserved, out-of-stock and older rows. They do not measure live on-demand prices or availability (the daily price history table above does that) and cannot establish savings against the current offers. Each day averages its recorded catalog snapshots; the median gives each day equal weight. Missing days are excluded.

Lowest daily catalog average$0.89/GPU/hr
Median of daily catalog averages$0.90/GPU/hr
Days with recorded catalog prices30 of 30
Latest catalog snapshotOctober 3, 2026

QuantaCloud

Need GPUs at scale?

Building out an inference fleet or training cluster? QuantaCloud brokers reserved capacity across multiple data center partners. 16+ GPUs, flexible terms, custom quote in 24 hours.

No waitlist24hr quote turnaroundInfiniBand fabric

Technical Specifications

CUDA cores
18176
Memory type
48 GB GDDR6
Tensor cores
568 (3rd Generation)
FP8 performance
365 TFLOPS
Memory bandwidth
864 GB/s

Strengths & Limitations

Advantages
  • High memory capacity (48GB) ideal for large datasets and complex models.
  • Excellent performance for both visualization and compute workloads.
  • Supports NVIDIA virtual GPU (vGPU) software for improved resource utilization and management.
  • Optimized for AI inference and training.
  • Strong FP32 and Tensor Core performance.
Limitations
  • Higher power consumption compared to lower-end GPUs.
  • May be overkill for less demanding tasks.
  • Price point is higher than entry-level data center GPUs.

Top Use Cases

AI Inference

The L40S excels at accelerating AI inference workloads, enabling real-time insights and decision-making. Its Tensor Cores provide significant performance gains for deep learning models.

Data Science and Analytics

With its large memory capacity and powerful compute capabilities, the L40S is well-suited for data science tasks such as data analysis, model training, and simulation.

Professional Visualization

The L40S delivers exceptional performance for professional visualization applications, including CAD, CAE, and digital content creation. It supports high-resolution displays and complex 3D models.

Real-World Benchmark

AI Inference Performance (ResNet-50)
The NVIDIA L40S demonstrates strong performance in AI inference benchmarks, such as ResNet-50, achieving high throughput and low latency. This makes it suitable for real-time AI applications.
Illustrative cost, generated estimate$0.22/hr

Market Analysis

Generated market commentary, not current pricing. Live prices and stock are in the sections above.

The NVIDIA L40S occupies a mid-range position in the data center GPU market, offering a compelling balance of performance and features for a variety of workloads. Its large memory capacity makes it a strong contender for memory-intensive applications.

Frequently Asked Questions

What is the typical power consumption of the NVIDIA L40S?▾

The typical power consumption of the NVIDIA L40S is around 300W.

Does the NVIDIA L40S support NVIDIA vGPU?▾

Yes, the NVIDIA L40S supports NVIDIA vGPU software, allowing for virtualization and sharing of GPU resources across multiple virtual machines.

What type of workloads is the NVIDIA L40S best suited for?▾

The NVIDIA L40S is well-suited for a wide range of workloads, including AI inference, data science, professional visualization, and virtual workstations.

Alternative GPUs

NVIDIA A30

Similar price point, but may offer different performance characteristics depending on the specific workload. The A30 has less memory (24GB) but may have better performance per dollar for certain compute tasks.

NVIDIA L4

A lower-cost alternative for less demanding workloads. The L4 has significantly lower memory (24GB) and compute performance, but is more power-efficient.

NVIDIA A10

A slightly less expensive option that still provides good performance for visualization and some AI tasks, although with less memory (24GB) and lower overall compute power.

Cite This Data
This pricing data is updated daily and free to cite with attribution.
Source: GPUPerHour.com — NVIDIA L40S GPU Rental Pricing Comparison (October 2026)

Journalists, bloggers, and researchers: You're welcome to cite our data in your articles with attribution. Our pricing database is checked periodically from 19 cloud providers tracking this GPU.

L40S Pricing: What It Costs in 2026

▾

L40S on-demand pricing currently ranges from $0.97/GPU/hr on Massed Compute to $1.55/GPU/hr on Ori, based on 12 offers reported in stock by 5 providers with live stock tracking, across 7 regions. GPUPerHour also tracks 119 on-demand listings in total, including catalog-only and out-of-stock rows, shown further up as last-seen pricing.

Running one L40S continuously for a month at the cheapest live rate costs approximately $698. Most providers bill per second or per minute, so shorter jobs cost proportionally less. Confirm the final price and capacity with the provider before booking.

Renting L40S: Which Provider to Choose

▾

5 providers have L40S in stock on-demand right now. The cheapest is Massed Compute at $0.97/GPU/hr, followed by RunPod at $1.09/GPU/hr and QuantaCloud at $1.09/GPU/hr.

Price is not the only factor when choosing a provider. Billing increments, region coverage, and security certifications also vary between providers. Use the pricing tool to filter by region, availability, and provider features.

How L40S Compares

▾

Compared to alternatives, the NVIDIA A30 offers another hardware option; follow the link for current prices. Similar price point, but may offer different performance characteristics depending on the specific workload. The A30 has less memory (24GB) but may have better performance per dollar for certain compute tasks. The NVIDIA L4 offers another hardware option; follow the link for current prices. A lower-cost alternative for less demanding workloads. The L4 has significantly lower memory (24GB) and compute performance, but is more power-efficient. The NVIDIA A10 offers another hardware option; follow the link for current prices. A slightly less expensive option that still provides good performance for visualization and some AI tasks, although with less memory (24GB) and lower overall compute power.

L40S Pricing FAQ

How much does L40S cost per hour?▾

L40S on-demand cloud rental currently starts at $0.97 per GPU per hour on Massed Compute and goes up to $1.55/GPU/hr, based on 12 offers reported in stock by 5 providers with live stock tracking. Running one L40S continuously for a month at the cheapest live rate costs approximately $698. Current prices use secure on-demand offers, excluding peer-to-peer hosts. Availability is checked periodically and can change before booking.

Which is the cheapest provider for L40S?▾

Among providers with L40S in stock right now, the cheapest are Massed Compute at $0.97/GPU/hr, RunPod at $1.09/GPU/hr, QuantaCloud at $1.09/GPU/hr. 12 offers are currently in stock across 5 providers.

How has the L40S rental price changed recently?▾

Over the last 14 days the cheapest on-demand L40S price reported in stock ranged from $0.80 to $0.97 per GPU-hour, across up to 8 providers with stock. The daily figures are captured at 00:05 UTC from secure on-demand offers and are listed in the price history table on this page.

What are the alternatives to L40S?▾

Alternatives to the L40S include NVIDIA A30, NVIDIA L4, NVIDIA A10. Similar price point, but may offer different performance characteristics depending on the specific workload. The A30 has less memory (24GB) but may have better performance per dollar for certain compute tasks.