RTX 2080 vs RTX 3060

TuringvsAmpereUpdated 6 days ago

Rent the RTX 3060 for most jobs. Ask for the 12 GB configuration, which holds a larger model or a longer context than the 8 GB most RTX 2080 cards carry, and it is a newer Ampere part with tensor cores that draws 170W against 215W. Pick the RTX 2080 only when your model already fits in 8 GB and the job is bandwidth bound, where its 448 GB/s is about 24 percent above the 360 GB/s of the RTX 3060.

RTX 2080 listed from $0.12/GPU/hrRTX 3060 listed from $0.27/GPU/hr

Right now, from live stock

  • Cheapest right now: RTX 3060 at $0.27/hr on Vast.ai

    Deploy
  • Most providers in stock: RTX 3060 (1)

    See all offers

Specifications Compared

SpecRTX 2080RTX 3060
TDP215W170W
VRAM8-11 GB8-12 GB
CUDA Cores2,9443,584
Memory TypeGDDR6GDDR6
ArchitectureTuringAmpere
FP16 (dense)40.3 TFLOPSNot published
Form FactorsPCIePCIe
INT8 (dense)161.1 TOPSNot published
InterconnectNVLink, PCIe 3.0PCIe 4.0
Tensor Cores368112
FP32 Performance10 TFLOPS12.7 TFLOPS
Memory Bandwidth448 GB/s360 GB/s

Performance Analysis

The 40.3 dense TFLOPS FP16 rating of the RTX 2080 enables higher throughput in mixed precision training loops than the RTX 3060 can demonstrate from its published 12.7 FP32 TFLOPS alone. Inference tasks that rely on dense INT8 operations benefit from the 161.1 dense TOPS figure available only on the RTX 2080. Memory bandwidth of 448 GB/s on the RTX 2080 supports larger batch sizes in memory bound kernels compared with the 360 GB/s limit of the RTX 3060. The 10 TFLOPS FP32 of the RTX 2080 versus 12.7 TFLOPS FP32 of the RTX 3060 indicates the newer card holds an edge in single precision workloads while the 45W lower TDP of the RTX 3060 reduces system power draw during sustained FP32 execution.

Current On-Demand Offers

Cheapest secure on-demand offer per provider that is reported in stock right now, per GPU per hour; multi-GPU instances show the whole-instance price alongside. Deploy opens the DeployGPU console for providers on the platform, otherwise the provider's own site. Stock is rechecked every minute.

RTX 2080

RTX 2080 is not offered on-demand by any provider we track right now. See the RTX 2080 rental page for last-seen listed prices and a price alert.

RTX 3060

ProviderRegionGPUsPer GPU / hrInstance / hrDeploy
Vast.aiCzechia, CZ4$0.27$1.08Deploy

1 provider in stock, 1 offer. All RTX 3060 offers, price history and alerts

Which GPU to watchWatch the price of

Notify me when RTX 2080 drops below a price

One email when the cheapest in-stock on-demand price per GPU-hour falls below your threshold.

QuantaCloud

Comparing providers? We broker across all of them.

Stop tab-switching between pricing pages. Tell us what you need, 16+ GPUs reserved or cluster capacity, and we return one quote at partner rates within 24 hours.

No waitlist24hr quote turnaroundInfiniBand fabric

Compare real-time pricing across 25+ providers

When to Choose the RTX 2080

Choose the RTX 2080 when your model and batch fit inside 8 GB and the job is bandwidth bound, since 448 GB/s beats the 360 GB/s of the RTX 3060. Its tensor cores are rated at 40.3 TFLOPS dense FP16, so small model FP16 inference runs well. Do not pick it for anything that needs more than 8 GB, and treat the NVLink listing as a footnote rather than a reason to rent it.

When to Choose the RTX 3060

Choose the RTX 3060 for most rentals. The 12 GB configuration loads models that an 8 GB RTX 2080 cannot, and its 112 Ampere tensor cores run FP16 and mixed precision work normally even though the table lists no figure. It also draws 170W against 215W. The only thing you give up is bandwidth, 360 GB/s against 448 GB/s, which matters mainly for small models that already fit in 8 GB.

Use Cases

LLM Training
RTX 2080

The RTX 2080 supplies 40.3 dense TFLOPS FP16 that the RTX 3060 does not publish.

LLM Inference
RTX 3060

The RTX 3060 delivers 12.7 FP32 TFLOPS at 170W TDP for standard inference paths.

Fine-tuning
RTX 2080

The RTX 2080 provides 161.1 dense TOPS INT8 and 448 GB/s bandwidth for fine tuning kernels.

Stable Diffusion
RTX 3060

Scientific Computing
RTX 3060

The RTX 3060 reaches 12.7 FP32 TFLOPS with 45W lower power than the RTX 2080.

Frequently Asked Questions

What FP32 performance does each GPU provide?

The RTX 2080 lists 10 TFLOPS FP32 while the RTX 3060 lists 12.7 TFLOPS FP32.

How does memory bandwidth differ between the cards?

The RTX 2080 reaches 448 GB/s bandwidth whereas the RTX 3060 reaches 360 GB/s bandwidth.

Which GPU publishes an FP16 rating?

The spec table lists 40.3 TFLOPS dense FP16 for the RTX 2080 and no dense FP16 figure for the RTX 3060. That is a gap in the published data, not a missing feature: the RTX 3060 has 112 tensor cores and runs FP16 and mixed precision workloads normally.

What TDP values apply to each model?

The RTX 2080 carries a 215W TDP and the RTX 3060 carries a 170W TDP.

Do both GPUs support the same interconnect?

The RTX 2080 includes NVLink while the RTX 3060 uses PCIe 4.0.

Which is cheaper to rent, the RTX 2080 or the RTX 3060?

Cloud rental prices for both the RTX 2080 and RTX 3060 vary by provider, configuration, and availability. This page shows live pricing from 25+ providers updated every 60 seconds. Scroll to the Live Cloud Pricing section to compare current rates.

How much VRAM does the RTX 2080 have compared to the RTX 3060?

The RTX 2080 has 8 to 11 GB of GDDR6 memory. The RTX 3060 has 8 to 12 GB of GDDR6 memory.

Can I find RTX 2080 and RTX 3060 GPUs available to rent right now?

Yes. This page shows real-time availability across 25+ cloud GPU providers. The Live Cloud Pricing section displays only in-stock offers with current pricing.

What is the main difference between the RTX 2080 and the RTX 3060?

The RTX 2080 uses the Turing architecture (2018) while the RTX 3060 uses Ampere (2021). The RTX 2080 has 8-11 GB of GDDR6 at 448 GB/s; the RTX 3060 has 8-12 GB of GDDR6 at 360 GB/s, so the RTX 2080 has 1.2x the memory bandwidth of the RTX 3060.

Rent these GPUs

Each GPU page lists every current on-demand offer by provider, per GPU-hour, updated every minute.

Related comparisons

How this page is made

  • Specifications come from the NVIDIA, AMD and Intel datasheets for the RTX 2080 and the RTX 3060. Dense and sparse throughput are listed separately.
  • Prices and availability are live: every offer shown is an in-stock, on-demand listing from a provider we track, rechecked every minute.
  • The written comparison (overview, performance notes, when to choose each card, verdict, use cases and FAQ) was drafted with an AI model from the specification table above and passed an automated check that rejects any figure not in that table. It was last generated on . It contains no prices; those are always read live.
  • Read how we collect the data, or report an error on this page.

Next steps