GPU comparison

B300 vs RTX 5090

Specifications and current cloud pricing, side by side.

Blackwell UltravsBlackwellUpdated 17 days ago

For the most common use case of LLM training the B300 emerges as the winner because its FP16 dense performance reaches 2250 TFLOPS and its memory capacity reaches 288 GB while the RTX 5090 remains limited to 32 GB and 209.5 TFLOPS under the same dense FP16 metric.

B300 from $7.89/GPU/hrRTX 5090 from $0.53/GPU/hr

Right now, from live stock

  • Cheapest right now: RTX 5090 at $0.53/hr on Vast.ai

    Deploy
  • Most providers in stock: B300 and RTX 5090 (3 each)

    See all offers
  • Best $ per TFLOPS FP16 dense right now: RTX 5090 ($0.0025 per TFLOPS-hour at 209.5 TFLOPS)

    Deploy

Specifications Compared

SpecB300RTX 5090
TDP1400W575W
VRAM262-288 GB32 GB
CUDA CoresNot published21,760
FP4 (dense)13,500 TFLOPS1,676 TFLOPS
FP8 (dense)4,500 TFLOPS419 TFLOPS
Memory TypeHBM3eGDDR7
ArchitectureBlackwell UltraBlackwell
FP16 (dense)2,250 TFLOPS209.5 TFLOPS
Form FactorsSXM6PCIe
INT8 (dense)187.5 TOPS838 TOPS
InterconnectNVLink, PCIe 6.0, InfiniBandPCIe 5.0
Tensor CoresNot published680
FP32 Performance75 TFLOPS104.8 TFLOPS
FP64 Performance1.25 TFLOPS1.6 TFLOPS
Memory Bandwidth8,000 GB/s1,792 GB/s
FP4 (with sparsity)18,000 TFLOPS3,352 TFLOPS
FP8 (with sparsity)9,000 TFLOPS838 TFLOPS
FP16 (with sparsity)4,500 TFLOPS419 TFLOPS
INT8 (with sparsity)375 TOPS1,676 TOPS

Performance Analysis

The FP16 dense performance of the B300 at 2250 TFLOPS exceeds the FP16 dense performance of the RTX 5090 at 209.5 TFLOPS by a ratio of 10.7 to 1. This ratio indicates substantially higher throughput for training and inference workloads that rely on dense FP16 arithmetic. The FP16 performance with sparsity of the B300 at 4500 TFLOPS likewise exceeds the FP16 performance with sparsity of the RTX 5090 at 419 TFLOPS by a ratio of 10.7 to 1. Memory bandwidth of 8000 GB/s on the B300 compared with 1792 GB/s on the RTX 5090 permits larger batch sizes during memory bound operations. The FP32 performance of the RTX 5090 at 104.8 TFLOPS exceeds the FP32 performance of the B300 at 75 TFLOPS.

Current On-Demand Offers

Cheapest secure on-demand offer per provider that is reported in stock right now, per GPU per hour; multi-GPU instances show the whole-instance price alongside. Deploy opens the DeployGPU console for providers on the platform, otherwise the provider's own site. Stock is rechecked every minute.

B300

ProviderRegionGPUsPer GPU / hrInstance / hrDeploy
RunPodglobal1$7.89—Deploy
LyceumEurope1$7.99—Deploy
VERDAFIN-HEL1$8.79—Deploy

3 providers in stock, 3 offers. All B300 offers, price history and alerts

RTX 5090

ProviderRegionGPUsPer GPU / hrInstance / hrDeploy
Vast.aiSouth Korea, KR1$0.53—Deploy
RunPodglobal1$0.99—Deploy
LeaderGPUThe Netherlands1$2.29—Deploy

3 providers in stock, 11 offers (cheapest per provider shown). All RTX 5090 offers, price history and alerts

Which GPU to watchWatch the price of

Notify me when B300 drops below a price

One email when the cheapest in-stock on-demand price per GPU-hour falls below your threshold. Cheapest right now: $7.89/GPU-hr.

QuantaCloud

Comparing B-series options? Get one quote for all of them.

Skip the per-provider sales calls. Reserved and cluster B-series configurations from 16 to 1024+ GPUs with InfiniBand fabric, 3 to 12 month terms. One quote at partner rates, 24h turnaround.

No waitlist24hr quote turnaroundInfiniBand fabric

Compare real-time pricing across 25+ providers

When to Choose the B300

The B300 constitutes the better choice for deployments that require VRAM capacity between 262 and 288 GB along with 8000 GB/s bandwidth. Its FP8 dense performance of 4500 TFLOPS and support for NVLink interconnect further suit large scale multi GPU configurations that exceed the capabilities of a single RTX 5090.

When to Choose the RTX 5090

The RTX 5090 constitutes the better choice for installations limited to a TDP of 575W and a PCIe form factor. Its FP32 performance of 104.8 TFLOPS supplies an advantage in workloads that depend on single precision floating point calculations.

Use Cases

LLM Training
B300

The B300 supplies FP16 dense performance of 2250 TFLOPS and VRAM capacity up to 288 GB which accommodates training runs that exceed the 32 GB limit of the RTX 5090.

LLM Inference
B300

The B300 supplies FP8 dense performance of 4500 TFLOPS together with 8000 GB/s bandwidth which supports higher throughput inference compared with the RTX 5090 FP8 dense performance of 419 TFLOPS.

Fine-tuning
B300

The B300 supplies FP16 dense performance of 2250 TFLOPS and memory bandwidth of 8000 GB/s which enable fine tuning of models that require capacity beyond 32 GB.

Stable Diffusion
RTX 5090

The RTX 5090 supplies FP32 performance of 104.8 TFLOPS at a TDP of 575W which aligns with desktop generation workloads that do not require the 1400W TDP of the B300.

Scientific Computing
Either

The B300 supplies higher FP16 dense performance of 2250 TFLOPS while the RTX 5090 supplies higher FP32 performance of 104.8 TFLOPS so selection depends on the required precision.

Frequently Asked Questions

What is the memory bandwidth difference between the B300 and the RTX 5090?▾

The B300 provides 8000 GB/s bandwidth while the RTX 5090 provides 1792 GB/s bandwidth.

How does FP16 dense performance compare between the two GPUs?▾

The B300 lists FP16 dense performance of 2250 TFLOPS while the RTX 5090 lists FP16 dense performance of 209.5 TFLOPS.

What TDP values are specified for each GPU?▾

The B300 lists a TDP of 1400W while the RTX 5090 lists a TDP of 575W.

Which interconnect options does each GPU support?▾

The B300 supports NVLink together with PCIe 6.0 and InfiniBand while the RTX 5090 supports PCIe 5.0.

How does FP32 performance differ between the B300 and the RTX 5090?▾

The B300 lists FP32 performance of 75 TFLOPS while the RTX 5090 lists FP32 performance of 104.8 TFLOPS.

Which is cheaper to rent, the B300 or the RTX 5090?▾

Cloud rental prices for both the B300 and RTX 5090 vary by provider, configuration, and availability. This page shows live pricing from 25+ providers updated every 60 seconds. Scroll to the Live Cloud Pricing section to compare current rates.

How much VRAM does the B300 have compared to the RTX 5090?▾

The B300 has 262 to 288 GB of HBM3e memory. The RTX 5090 has 32 GB of GDDR7 memory.

Can I find B300 and RTX 5090 GPUs available to rent right now?▾

Yes. This page shows real-time availability across 25+ cloud GPU providers. The Live Cloud Pricing section displays only in-stock offers with current pricing.

What is the main difference between the B300 and the RTX 5090?▾

The B300 uses the Blackwell Ultra architecture (2025) while the RTX 5090 uses Blackwell (2025). The B300 delivers 10.7x the dense FP16 throughput (2,250 vs 209.5 TFLOPS, both without sparsity) and 4.5x the memory bandwidth of the RTX 5090.

Rent these GPUs

Each GPU page lists every current on-demand offer by provider, per GPU-hour, updated every minute.

Related comparisons

How this page is made

  • Specifications come from the NVIDIA, AMD and Intel datasheets for the B300 and the RTX 5090. Dense and sparse throughput are listed separately.
  • Prices and availability are live: every offer shown is an in-stock, on-demand listing from a provider we track, rechecked every minute.
  • The written comparison (overview, performance notes, when to choose each card, verdict, use cases and FAQ) was drafted with an AI model from the specification table above and passed an automated check that rejects any figure not in that table. It was last generated on . It contains no prices; those are always read live.
  • Read how we collect the data, or report an error on this page.

Next steps