$ AIPriceBoardbeta

B200-180GB GPUs

Dedicated B200-180GB instances rented per hour: SSH-rentable containers and dedicated LLM deployment pods. Prices are synced from provider APIs.

B200-180GB GPUs rental offers

ConfigurationGPU$/hourStatusKindProviderLast sync
11 1xB200-180GB logo 1xB200-180GB B200-180GB $3.69 Available LLM deploy DeepInfra
11 1xB200-180GB logo 1xB200-180GB B200-180GB $3.69 Busy SSH rental DeepInfra
21 2xB200-180GB logo 2xB200-180GB B200-180GB $7.38 Available LLM deploy DeepInfra
21 2xB200-180GB logo 2xB200-180GB B200-180GB $7.38 Busy SSH rental DeepInfra
41 4xB200-180GB logo 4xB200-180GB B200-180GB $14.76 Busy SSH rental DeepInfra
41 4xB200-180GB logo 4xB200-180GB B200-180GB $14.76 Busy LLM deploy DeepInfra
81 8xB200-180GB logo 8xB200-180GB B200-180GB $29.52 Busy LLM deploy DeepInfra
81 8xB200-180GB logo 8xB200-180GB B200-180GB $29.52 Busy SSH rental DeepInfra

Availability and prices reflect the last sync (14h ago). Hourly billing, no token charges on dedicated hardware.

Frequently asked questions

How is dedicated GPU rental billed?

Per hour of instance uptime, in USD. Unlike serverless inference there are no token charges — you rent the whole card or pod and can run any workload on it, 24/7 or spot.

What is the difference between an LLM deployment pod and an SSH container?

A pod comes with a model-serving stack preinstalled, so you can deploy and scale an LLM in minutes. An SSH container is a plain machine you fully control: install anything, train, fine-tune or serve multiple models.

Where do these prices come from?

Straight from each provider's public infrastructure API during the nightly sync. The "last checked" timestamp on every row shows when that offer was last verified.