Vast.ai48GB VRAMAda Lovelaceenterprise

L40 on Vast.ai

Visit Vast.ai

Vast.ai provides access to the NVIDIA L40 GPU, equipped with 48GB GDDR6 VRAM on the Ada Lovelace architecture, optimized for enterprise data center workloads including AI inference, visualization, and rendering. This decentralized marketplace stands out by offering the absolute lowest rental costs—frequently under $1 per hour—through direct host competition, making high-VRAM GPUs accessible without enterprise contracts. Noteworthy for ML engineers and data scientists handling large-scale inference on models like Llama 70B or Stable Diffusion, it excels in cost-per-performance via granular filters such as DLPerf/$. Key value propositions include per-second billing, spot instances for up to 70% savings, and support for distributed experiments across global hosts. Ideal for budget-conscious teams prioritizing ROI on memory-intensive AI tasks over premium support.

Why NVIDIA L40 on Vast.ai?

Vast.ai paired with NVIDIA L40 offers unmatched cost efficiency for its 48GB VRAM and inference prowess, leveraging the decentralized marketplace to undercut traditional clouds by 50-80%. Hosts compete on price, enabling per-hour rates as low as $0.60-$1.20, with spot instances for interruptible workloads slashing costs further. Granular filters like DLPerf/$, VRAM/$, and host reliability ensure optimal selection for L40's strengths in FP8/FP16 inference and ray tracing. This combo suits short-term experiments or scale-outs, complementing L40's enterprise tier with flexible scaling, no egress fees, and Docker-based deployments for rapid PyTorch/TensorFlow setups—perfect for cost-sensitive ML teams avoiding lock-in.

Live Pricing

Real-time NVIDIA L40 offers from Vast.ai

1 of 50 in stock
Vast.ai
Vast.ai
Slovenia
Available
NVIDIA L40S
48GB VRAM
256 vCPU
189GB RAM
2938GB Storage
5281 Mbps ↑
4628 Mbps ↓
$0.80/GPU/hr
Vast.ai
Vast.ai
Virginia
Sold Out
NVIDIA L40S
48GB VRAM
8 vCPU
62GB RAM
195GB Storage
2391 Mbps ↑
9568 Mbps ↓
$0.03/GPU/hr
Vast.ai
Vast.ai
Belarus
Sold Out
NVIDIA L403x
48GB VRAM
192 vCPU
755GB RAM
2944GB Storage
281 Mbps ↑
527 Mbps ↓
$0.33/GPU/hr
$1.00/hr total (3×)
Vast.ai
Vast.ai
Vietnam
Sold Out
NVIDIA L40
48GB VRAM
96 vCPU
188GB RAM
718GB Storage
404 Mbps ↑
684 Mbps ↓
$0.33/GPU/hr
Vast.ai
Vast.ai
Vietnam
Sold Out
NVIDIA L40
48GB VRAM
96 vCPU
125GB RAM
334GB Storage
423 Mbps ↑
670 Mbps ↓
$0.33/GPU/hr

Performance Notes

NVIDIA L40 on Vast.ai delivers robust Ada Lovelace performance: ~90 TFLOPS FP16, 181 TFLOPS sparse Tensor FP16, ideal for inference on 30-70B models with 48GB VRAM. Host-dependent factors include 10-100Gbps networking (adequate for most distributed jobs), NVMe SSD storage (1-20TB typical), and PCIe 4.0/NVLink for multi-GPU scaling up to 8x with near-linear efficiency in NCCL benchmarks. DLPerf scores (via Vast.ai filters) indicate reliable ML throughput, but variability exists due to host configs (CUDA 12+, driver 535+). Spot instances risk preemption; on-demand offers stability. Unknowns: exact inter-host InfiniBand prevalence—verify per listing for H100-like scaling.

About Vast.ai

A decentralized marketplace for absolute lowest costs and distributed experiments.

Best For

Absolute lowest costsDistributed experiments

Unique Features

  • Granular search filters like DLPerf/$
  • Decentralized marketplace
NVIDIA L40 Specs

VRAM

48GB

Architecture

Ada Lovelace

Tier

enterprise

Platform Features

Access Methods
SSH
Jupyter Notebooks
Web Terminal
API
Kubernetes
Containers
Billing Options
Incrementper-hour
Spot Instances
Reserved Instances
Prepaid Credits
Compliance
SOC 2
HIPAA
GDPR
ISO 27001

Getting Started

Launch NVIDIA L40 on Vast.ai quickly through its web dashboard: search verified hosts, filter by performance metrics, and deploy pre-configured ML images. Pay per second with no upfront commitments, supporting instant scaling for experiments. Focus on DLPerf/$ for value.

Steps

  1. 1Create Vast.ai account and deposit funds via card/crypto (minimum $5).
  2. 2Search 'NVIDIA L40', filter by DLPerf/$, uptime >99%, verified hosts.
  3. 3Select on-demand/spot, customize CPU/RAM (32+ cores/128GB rec.), disk size.
  4. 4Pick template (PyTorch 2.3, TensorFlow, Jupyter) and click 'Rent'.
  5. 5Connect via SSH/NoVNC; workloads start in <2 minutes.

Pro Tips

  • Sort by DLPerf/$ and test short rentals first to benchmark host-specific L40 perf.
  • Enable auto-relaunch on spot instances for fault-tolerant distributed training.
  • Use Vast.ai CLI for scripting multi-instance deployments across L40 clusters.

Frequently Asked Questions

What is Vast.ai's billing model for NVIDIA L40?

Vast.ai bills per-hour for GPU instances including NVIDIA L40. Hourly billing means you pay for full hours even if your job completes mid-hour. Plan your workloads accordingly to maximize cost efficiency.

Does Vast.ai offer spot instances for NVIDIA L40?

Yes, Vast.ai offers spot/preemptible instances for NVIDIA L40, which can reduce costs by 50-80% compared to on-demand pricing. Spot instances are ideal for fault-tolerant workloads like batch inference, hyperparameter tuning, and training jobs with checkpointing. Note that spot instances can be interrupted when demand is high, so ensure your workflow can handle preemption gracefully.

Related Pages

Compare L40 Across Providers

The L40 is available from 15 providers on GPUPerHour. Vast.ai charges $0.80/hr. Here is how other providers compare:

For a full comparison across all providers, see the L40 rental page. See all GPUs on Vast.ai.