$ AIPriceBoardbeta

GB200 NVL72-186GB GPUs

Dedicated GB200 NVL72-186GB instances rented per hour: SSH-rentable containers and dedicated LLM deployment pods. Prices are synced from provider APIs.

GB200 NVL72-186GB GPUs rental offers

ConfigurationGPU$/hourStatusKindProviderLast sync
4N 4xGB200 NVL72-186GB logo 4xGB200 NVL72-186GB GB200 NVL72-186GB $42 Available LLM deploy CoreWeave

Availability and prices reflect the last sync (50m ago). Hourly billing, no token charges on dedicated hardware.

Frequently asked questions

How is dedicated GPU rental billed?

Per hour of instance uptime, in USD. Unlike serverless inference there are no token charges — you rent the whole card or pod and can run any workload on it, 24/7 or spot.

What is the difference between an LLM deployment pod and an SSH container?

A pod comes with a model-serving stack preinstalled, so you can deploy and scale an LLM in minutes. An SSH container is a plain machine you fully control: install anything, train, fine-tune or serve multiple models.

Where do these prices come from?

Straight from each provider's public infrastructure API during the nightly sync. The "last checked" timestamp on every row shows when that offer was last verified.