A40 vs RTX A6000

AmperevsAmpereUpdated yesterday

The RTX A6000 wins for the most common use case of mixed precision training and inference because its 154.8 TFLOPS dense FP16 rating, 768 GB/s bandwidth, and 38.7 TFLOPS FP32 rating exceed the corresponding A40 figures while both cards share identical 48 GB memory capacity and 300W TDP.

A40 from $0.07/GPU/hrRTX A6000 from $0.44/GPU/hr

Right now, from live stock

  • Cheapest right now: A40 at $0.07/hr on Vultr

    Deploy
  • Most providers in stock: RTX A6000 (7)

    See all offers
  • Best $ per TFLOPS FP16 dense right now: A40 ($0.0005 per TFLOPS-hour at 149.7 TFLOPS)

    Deploy

Specifications Compared

SpecA40RTX A6000
TDP300W300W
VRAM48 GB48 GB
CUDA Cores10,75210,752
Memory TypeGDDR6GDDR6
ArchitectureAmpereAmpere
FP16 (dense)149.7 TFLOPS154.8 TFLOPS
Form FactorsPCIePCIe
INT8 (dense)299.3 TOPS309.7 TOPS
InterconnectNVLink, PCIe 4.0NVLink, PCIe 4.0
Tensor Cores336336
FP32 Performance37.4 TFLOPS38.7 TFLOPS
Memory Bandwidth696 GB/s768 GB/s
FP16 (with sparsity)299.4 TFLOPS309.6 TFLOPS
INT8 (with sparsity)598.6 TOPS619.4 TOPS

Performance Analysis

The ratio of dense FP16 to FP32 equals 4.0 on the A40 with its 149.7 TFLOPS dense FP16 figure and 4.0 on the RTX A6000 with its 154.8 TFLOPS dense FP16 figure, indicating that both cards accelerate mixed precision workloads by the same factor during training and inference. The RTX A6000 further widens its advantage when sparsity is enabled because its 309.6 TFLOPS with sparsity exceeds the A40 value of 299.4 TFLOPS with sparsity by 10.2 TFLOPS. Memory bandwidth of 768 GB/s on the RTX A6000 versus 696 GB/s on the A40 permits larger batch sizes in memory bound kernels while the INT8 dense rating of 309.7 TOPS on the RTX A6000 exceeds the 299.3 TOPS dense rating on the A40.

Current On-Demand Offers

Cheapest secure on-demand offer per provider that is reported in stock right now, per GPU per hour; multi-GPU instances show the whole-instance price alongside. Deploy opens the DeployGPU console for providers on the platform, otherwise the provider's own site. Stock is rechecked every minute.

A40

ProviderRegionGPUsPer GPU / hrInstance / hrDeploy
VultrNew Jersey, US1$0.07Deploy
RunPodglobal1$0.49Deploy
LeaderGPUThe Netherlands8$0.52$4.13Deploy

3 providers in stock, 3 offers. All A40 offers, price history and alerts

RTX A6000

ProviderRegionGPUsPer GPU / hrInstance / hrDeploy
LeaderGPUThe Netherlands8$0.44$3.54Deploy
QuantaCloudus-midwest-21$0.48Deploy
HyperstackCANADA-11$0.50Deploy
RunPodglobal1$0.53Deploy
Massed Computeus-central-22$0.55$1.10Deploy

7 providers in stock, 58 offers (cheapest per provider shown). All RTX A6000 offers, price history and alerts

Which GPU to watchWatch the price of

Notify me when A40 drops below a price

One email when the cheapest in-stock on-demand price per GPU-hour falls below your threshold. Cheapest right now: $0.07/GPU-hr.

QuantaCloud

Comparing providers? We broker across all of them.

Stop tab-switching between pricing pages. Tell us what you need, 16+ GPUs reserved or cluster capacity, and we return one quote at partner rates within 24 hours.

No waitlist24hr quote turnaroundInfiniBand fabric

Compare real-time pricing across 25+ providers

When to Choose the A40

The A40 suits deployments that prioritize the lower dense FP16 rating of 149.7 TFLOPS when paired with the 696 GB/s bandwidth figure in environments where the 5.1 TFLOPS dense FP16 gap versus the RTX A6000 remains acceptable. Its 37.4 TFLOPS FP32 rating also matches workloads that do not require the additional 1.3 TFLOPS available on the RTX A6000.

When to Choose the RTX A6000

The RTX A6000 suits deployments that require the higher memory bandwidth of 768 GB/s to support increased batch sizes alongside its 154.8 TFLOPS dense FP16 rating. Its 38.7 TFLOPS FP32 rating and 309.7 TOPS INT8 dense rating further favor selection when the 1.3 TFLOPS and 10.4 TOPS margins over the A40 improve throughput in scientific and inference pipelines.

Use Cases

LLM Training
RTX A6000

The RTX A6000 supplies 154.8 TFLOPS dense FP16 compared with 149.7 TFLOPS on the A40.

LLM Inference
RTX A6000

The RTX A6000 supplies 309.7 TOPS INT8 dense compared with 299.3 TOPS on the A40.

Fine-tuning
Either

Both cards provide identical 48 GB GDDR6 memory capacity for parameter storage.

Stable Diffusion
RTX A6000

The RTX A6000 supplies 768 GB/s bandwidth compared with 696 GB/s on the A40.

Scientific Computing
RTX A6000

The RTX A6000 supplies 38.7 TFLOPS FP32 compared with 37.4 TFLOPS on the A40.

Frequently Asked Questions

What is the memory bandwidth difference between the A40 and RTX A6000?

The A40 provides 696 GB/s while the RTX A6000 provides 768 GB/s.

What dense FP16 performance do these GPUs deliver?

The A40 delivers 149.7 TFLOPS dense FP16 while the RTX A6000 delivers 154.8 TFLOPS dense FP16.

Do the A40 and RTX A6000 share the same TDP?

Both cards list a TDP of 300W.

Which GPU offers higher INT8 performance with sparsity?

The RTX A6000 offers 619.4 TOPS INT8 with sparsity while the A40 offers 598.6 TOPS INT8 with sparsity.

What FP32 ratings appear for the A40 and RTX A6000?

The A40 lists 37.4 TFLOPS FP32 while the RTX A6000 lists 38.7 TFLOPS FP32.

Which is cheaper to rent, the A40 or the RTX A6000?

Cloud rental prices for both the A40 and RTX A6000 vary by provider, configuration, and availability. This page shows live pricing from 25+ providers updated every 60 seconds. Scroll to the Live Cloud Pricing section to compare current rates.

How much VRAM does the A40 have compared to the RTX A6000?

The A40 has 48 GB of GDDR6 memory. The RTX A6000 has 48 GB of GDDR6 memory.

Can I find A40 and RTX A6000 GPUs available to rent right now?

Yes. This page shows real-time availability across 25+ cloud GPU providers. The Live Cloud Pricing section displays only in-stock offers with current pricing.

What is the main difference between the A40 and the RTX A6000?

The A40 uses the Ampere architecture (2020) while the RTX A6000 uses Ampere (2020). Both deliver the same dense FP16 throughput (154.8 TFLOPS without sparsity), and the RTX A6000 has 1.1x the memory bandwidth of the A40.

Rent these GPUs

Each GPU page lists every current on-demand offer by provider, per GPU-hour, updated every minute.

Related comparisons

How this page is made

  • Specifications come from the NVIDIA, AMD and Intel datasheets for the A40 and the RTX A6000. Dense and sparse throughput are listed separately.
  • Prices and availability are live: every offer shown is an in-stock, on-demand listing from a provider we track, rechecked every minute.
  • The written comparison (overview, performance notes, when to choose each card, verdict, use cases and FAQ) was drafted with an AI model from the specification table above and passed an automated check that rejects any figure not in that table. It was last generated on . It contains no prices; those are always read live.
  • Read how we collect the data, or report an error on this page.

Next steps