$ AIPriceBoardbeta

H200-141GB GPUs

Dedicated H200-141GB instances rented per hour: SSH-rentable containers and dedicated LLM deployment pods. Prices are synced from provider APIs.

H200-141GB GPUs rental offers

ConfigurationGPU$/hourStatusKindProviderLast sync
11 1xH200-141GB logo 1xH200-141GB H200-141GB $2.68999992 Available LLM deploy DeepInfra
11 1xH200-141GB logo 1xH200-141GB H200-141GB $4.47 Available LLM deploy DigitalOcean
21 2xH200-141GB logo 2xH200-141GB H200-141GB $5.37999984 Available LLM deploy DeepInfra
41 4xH200-141GB logo 4xH200-141GB H200-141GB $10.75999968 Busy LLM deploy DeepInfra
81 8xH200-141GB logo 8xH200-141GB H200-141GB $21.51999936 Busy LLM deploy DeepInfra

Availability and prices reflect the last sync (14h ago). Hourly billing, no token charges on dedicated hardware.

Frequently asked questions

How is dedicated GPU rental billed?

Per hour of instance uptime, in USD. Unlike serverless inference there are no token charges — you rent the whole card or pod and can run any workload on it, 24/7 or spot.

What is the difference between an LLM deployment pod and an SSH container?

A pod comes with a model-serving stack preinstalled, so you can deploy and scale an LLM in minutes. An SSH container is a plain machine you fully control: install anything, train, fine-tune or serve multiple models.

Where do these prices come from?

Straight from each provider's public infrastructure API during the nightly sync. The "last checked" timestamp on every row shows when that offer was last verified.