$ AIPriceBoardbeta

H200 141 GB GPUs

Dedicated H200 141 GB instances rented per hour: SSH-rentable containers and dedicated LLM deployment pods. Prices are synced from provider APIs.

H200 141 GB GPUs rental offers

ConfigurationGPU$/hourStatusKindProviderLast sync
H1 H200 141 GB GPU logo H200 141 GB GPU H200 141 GB $8 Available LLM deploy Fireworks AI

Availability and prices reflect the last sync (53m ago). Hourly billing, no token charges on dedicated hardware.

Frequently asked questions

How is dedicated GPU rental billed?

Per hour of instance uptime, in USD. Unlike serverless inference there are no token charges — you rent the whole card or pod and can run any workload on it, 24/7 or spot.

What is the difference between an LLM deployment pod and an SSH container?

A pod comes with a model-serving stack preinstalled, so you can deploy and scale an LLM in minutes. An SSH container is a plain machine you fully control: install anything, train, fine-tune or serve multiple models.

Where do these prices come from?

Straight from each provider's public infrastructure API during the nightly sync. The "last checked" timestamp on every row shows when that offer was last verified.