The cheapest B200 on the table right now is $1.81 per GPU-hour at Google Cloud (spot), out of 41 offers from 18 providers. The B200 carries 192 GB of VRAM on NVIDIA's Blackwell generation.
On-demand pricing for the same card spans $4.90 to $14.24 per GPU-hour — a 2.9× spread between the cheapest and the most expensive provider for identical silicon. That gap is the entire reason this table exists.
Cheapest by pricing model: on-demand $4.90 at GPU.ai · spot $1.81 at Google Cloud · reserved $3.49 at fal.
| Provider | Tier | $/GPU-hr | Pricing | Node | Commit | In stock | |
|---|---|---|---|---|---|---|---|
| Google Cloud | dedicated | $1.81 | spot | 1× GPU | — | — | rent |
| Verda | dedicated | $3.06 | spot | 1–8× GPU | — | — | rent |
| fal | dedicated | $3.49 | reserved | 1× GPU | — | — | rent |
| Nebius | dedicated | $3.95 | reserved | 8× GPU | — | — | rent |
| Vast.ai | marketplace | $4.38 | spot | 8× GPU | — | yes | rent |
| GPU.ai | dedicated | $4.90 | on-demand | 1× GPU | — | yes | rent |
| Google Cloud | dedicated | $4.95 | on-demand | 1× GPU | — | — | rent |
| Vast.ai | marketplace | $5.00 | spot | 1–4× GPU | — | yes | rent |
| AWS | dedicated | $5.20 | spot | 8× GPU | — | — | rent |
| Vast.ai | marketplace | $5.38 | on-demand | 8× GPU | — | yes | rent |
| Massed Compute | dedicated | $5.43 | on-demand | 8× GPU | — | — | rent |
| Beam | dedicated | $5.62 | on-demand | 1× GPU | — | yes | rent |
| Vast.ai | marketplace | $5.76 | on-demand | 2× GPU | — | yes | rent |
| Vast.ai | marketplace | $5.88 | on-demand | 1× GPU | — | yes | rent |
| Hyperstack | dedicated | $6.00 | on-demand | 1× GPU | — | — | rent |
| Vast.ai | marketplace | $6.00 | on-demand | 4× GPU | — | yes | rent |
| Verda | dedicated | $6.11 | on-demand | 1–8× GPU | — | — | rent |
| Modal | dedicated | $6.25 | on-demand | 1× GPU | — | yes | rent |
| fal | dedicated | $6.25 | on-demand | 1× GPU | — | — | rent |
| Lambda Labs | dedicated | $6.69 | on-demand | 8× GPU | — | — | rent |
| Runpod | dedicated | $6.79 | on-demand | 1× GPU | — | yes | rent |
| Runpod | dedicated | $6.79 | spot | 1× GPU | — | yes | rent |
| Lambda Labs | dedicated | $6.79 | on-demand | 4× GPU | — | — | rent |
| Lambda Labs | dedicated | $6.89 | on-demand | 2× GPU | — | — | rent |
| Lambda Labs | dedicated | $6.99 | on-demand | 1× GPU | — | — | rent |
| Nebius | dedicated | $7.15 | on-demand | 8× GPU | — | — | rent |
| Spheron | dedicated | $7.20 | on-demand | 1× GPU | — | — | rent |
| Together AI | dedicated | $7.99 | reserved | 8× GPU | — | — | rent |
| Together AI | dedicated | $8.19 | on-demand | 8× GPU | — | — | rent |
| CoreWeave | dedicated | $8.60 | on-demand | 8× GPU | — | — | rent |
| Baseten | dedicated | $9.98 | on-demand | 1× GPU | — | yes | rent |
| Oracle Cloud | dedicated | $14.00 | on-demand | 8× GPU | — | — | rent |
| AWS | dedicated | $14.24 | on-demand | 8× GPU | — | — | rent |
How much does a B200 cost per hour?
As of the latest scrape, $1.81 per GPU-hour at Google Cloud is the cheapest B200 offer across 18 tracked cloud providers. The cheapest on-demand rate is $4.90 per GPU-hour at GPU.ai, and the most expensive on-demand rate is $14.24 at AWS.
Which cloud is cheapest for the B200?
Google Cloud at $1.81 per GPU-hour on spot capacity. Prices move constantly, so this page is regenerated from the live scrape every 15 minutes. Note that marketplace tiers are peer hardware and are not directly comparable to dedicated capacity.
Is spot B200 capacity cheaper?
Yes — the cheapest spot B200 is $1.81 per GPU-hour at Google Cloud, -63% versus the cheapest on-demand rate of $4.90. Spot instances can be reclaimed by the provider at any time, so they suit checkpointed training and batch inference rather than long-lived services.
How many GB of VRAM does the B200 have?
192 GB, on the Blackwell architecture. Multi-GPU instances multiply that: an 8× B200 node exposes 1536 GB of GPU memory in total.