5 articles with this tag
VRAM, memory bandwidth and dense throughput decide an AI workload. How to find them on a datasheet, and why the headline TFLOPS figure is usually doubled.
Verified purchase listings for NVIDIA H200, B200, B300, GB200 NVL72 and A100 with seller and date, the misquoted figures corrected, and live rental prices.
What an H100 card, an 8-GPU server and a DGX H100 last listed for, with sellers and dates, next to live rental prices and a break-even method you can run.
NVLink, PCIe and SXM explained for people renting multi-GPU servers: bandwidth per generation, what the H100 NVL is, and when the SXM premium pays off.
MIG splits one NVIDIA data centre GPU into isolated instances with their own memory. See which GPUs support it and when a slice beats a whole cheap GPU.