12 articles in this category
Pick a precision for pre-training, fine-tuning or inference, and see which rentable GPUs accelerate FP8 and FP4 in hardware and which only emulate them.
Which free GPU options still work in September 2026, which closed, who qualifies for credits, and what the cheapest paid GPU costs once the quota runs out.
Colab plans and compute units explained, a dated estimate of what a Colab T4 costs per hour, Kaggle's free tier, and live hourly prices to rent the same GPUs.
Who rents cloud GPUs, grouped by kind and listed alphabetically, with a live price table, verified facts per provider and five questions to shortlist three.
How on-demand, spot, reserved, capacity block and marketplace GPU pricing work, what each one commits you to, and a rule for picking between them.
Pick between a monthly dedicated GPU server, a GPU VPS and an hourly cloud GPU for an always-on job, with dated list prices and a break-even formula.
VRAM, memory bandwidth and dense throughput decide an AI workload. How to find them on a datasheet, and why the headline TFLOPS figure is usually doubled.
A is Ampere, L is Ada, H is Hopper, B is Blackwell. A decoder table, a generation timeline and a four-step method for placing any NVIDIA GPU on a rental site.
NVLink, PCIe and SXM explained for people renting multi-GPU servers: bandwidth per generation, what the H100 NVL is, and when the SXM premium pays off.
A GPU-hour is one GPU used for one hour. Learn how providers meter it, what the price leaves out, and how to estimate a job's cost from live prices.
A neocloud is a GPU-first cloud. See how neoclouds differ from hyperscalers and GPU marketplaces, who they are, the risks, and which one to rent from.
MIG splits one NVIDIA data centre GPU into isolated instances with their own memory. See which GPUs support it and when a slice beats a whole cheap GPU.