L40 on Vast.ai
Visit Vast.aiVast.ai provides access to the NVIDIA L40 GPU, equipped with 48GB GDDR6 VRAM on the Ada Lovelace architecture, optimized for enterprise data center workloads including AI inference, visualization, and rendering. This decentralized marketplace stands out by offering the absolute lowest rental costs—frequently under $1 per hour—through direct host competition, making high-VRAM GPUs accessible without enterprise contracts. Noteworthy for ML engineers and data scientists handling large-scale inference on models like Llama 70B or Stable Diffusion, it excels in cost-per-performance via granular filters such as DLPerf/$. Key value propositions include per-second billing, spot instances for up to 70% savings, and support for distributed experiments across global hosts. Ideal for budget-conscious teams prioritizing ROI on memory-intensive AI tasks over premium support.
Why NVIDIA L40 on Vast.ai?
Vast.ai paired with NVIDIA L40 offers unmatched cost efficiency for its 48GB VRAM and inference prowess, leveraging the decentralized marketplace to undercut traditional clouds by 50-80%. Hosts compete on price, enabling per-hour rates as low as $0.60-$1.20, with spot instances for interruptible workloads slashing costs further. Granular filters like DLPerf/$, VRAM/$, and host reliability ensure optimal selection for L40's strengths in FP8/FP16 inference and ray tracing. This combo suits short-term experiments or scale-outs, complementing L40's enterprise tier with flexible scaling, no egress fees, and Docker-based deployments for rapid PyTorch/TensorFlow setups—perfect for cost-sensitive ML teams avoiding lock-in.
Live Pricing
Real-time NVIDIA L40 offers from Vast.ai
| Provider | GPU Model | VRAM | Host Specs | Region | Price | Status | Action | |
|---|---|---|---|---|---|---|---|---|
![]() Vast.ai | NVIDIA L40S 48GB VRAM | 48GB | 8 vCPU 62GB RAM 195GB Storage | Virginia | $0.03/GPU/hr | Sold Out | ||
![]() Vast.ai | NVIDIA L40 48GB VRAM | 48GB | 96 vCPU 188GB RAM 718GB Storage | Vietnam | $0.33/GPU/hr | Sold Out | ||
![]() Vast.ai | NVIDIA L40 48GB VRAM | 48GB | 96 vCPU 125GB RAM 334GB Storage | Vietnam | $0.33/GPU/hr | Sold Out | ||
![]() Vast.ai | 3×NVIDIA L40 48GB VRAM | 48GB | 192 vCPU 755GB RAM 2944GB Storage | Belarus | $0.33/GPU/hr $1.00/hr total (3×) | Sold Out | ||
![]() Vast.ai | NVIDIA L40S 48GB VRAM | 48GB | 256 vCPU 189GB RAM 2938GB Storage | Slovenia | $0.80/GPU/hr | Available |





Performance Notes
NVIDIA L40 on Vast.ai delivers robust Ada Lovelace performance: ~90 TFLOPS FP16, 181 TFLOPS sparse Tensor FP16, ideal for inference on 30-70B models with 48GB VRAM. Host-dependent factors include 10-100Gbps networking (adequate for most distributed jobs), NVMe SSD storage (1-20TB typical), and PCIe 4.0/NVLink for multi-GPU scaling up to 8x with near-linear efficiency in NCCL benchmarks. DLPerf scores (via Vast.ai filters) indicate reliable ML throughput, but variability exists due to host configs (CUDA 12+, driver 535+). Spot instances risk preemption; on-demand offers stability. Unknowns: exact inter-host InfiniBand prevalence—verify per listing for H100-like scaling.
A decentralized marketplace for absolute lowest costs and distributed experiments.
Best For
Unique Features
- Granular search filters like DLPerf/$
- Decentralized marketplace
VRAM
48GB
Architecture
Ada Lovelace
Tier
enterprise
Platform Features
Getting Started
Launch NVIDIA L40 on Vast.ai quickly through its web dashboard: search verified hosts, filter by performance metrics, and deploy pre-configured ML images. Pay per second with no upfront commitments, supporting instant scaling for experiments. Focus on DLPerf/$ for value.
Steps
- 1Create Vast.ai account and deposit funds via card/crypto (minimum $5).
- 2Search 'NVIDIA L40', filter by DLPerf/$, uptime >99%, verified hosts.
- 3Select on-demand/spot, customize CPU/RAM (32+ cores/128GB rec.), disk size.
- 4Pick template (PyTorch 2.3, TensorFlow, Jupyter) and click 'Rent'.
- 5Connect via SSH/NoVNC; workloads start in <2 minutes.
Pro Tips
- Sort by DLPerf/$ and test short rentals first to benchmark host-specific L40 perf.
- Enable auto-relaunch on spot instances for fault-tolerant distributed training.
- Use Vast.ai CLI for scripting multi-instance deployments across L40 clusters.
Frequently Asked Questions
What is Vast.ai's billing model for NVIDIA L40?▾
Vast.ai bills per-hour for GPU instances including NVIDIA L40. Hourly billing means you pay for full hours even if your job completes mid-hour. Plan your workloads accordingly to maximize cost efficiency.
Does Vast.ai offer spot instances for NVIDIA L40?▾
Yes, Vast.ai offers spot/preemptible instances for NVIDIA L40, which can reduce costs by 50-80% compared to on-demand pricing. Spot instances are ideal for fault-tolerant workloads like batch inference, hyperparameter tuning, and training jobs with checkpointing. Note that spot instances can be interrupted when demand is high, so ensure your workflow can handle preemption gracefully.
Related Pages
Rent NVIDIA L40
Atlantic.net vs Vast.ai: GPU Cloud Comparison
AWS vs Vast.ai: GPU Cloud Comparison
Cirrascale vs Vast.ai: GPU Cloud Comparison
NVIDIA A10 on Vast.ai - Pricing & Availability
NVIDIA A100 PCIe 80GB on Vast.ai - Pricing & Availability
NVIDIA A100 SXM4 40GB on Vast.ai - Pricing & Availability
NVIDIA A100 SXM4 80GB on Vast.ai - Pricing & Availability
NVIDIA L40 in Belarus - Pricing & Availability
NVIDIA L40 in Finland - Pricing & Availability