B200 and B300 Rental
What NVIDIA Blackwell GPUs Cost to Rent in 2026, On Demand and Reserved, and How to Get Allocation
B200s rent for a median $6.49 an hour and B300s about $7.99, but fewer than one listing in five is confirmed in stock. We find who can actually deliver and negotiate the terms.
Get B200 and B300 Quotes Through UsB200 and B300 rental is where the AI hardware market is moving in 2026. The NVIDIA B200 is the first Blackwell data center GPU, with 180 GB of memory and roughly two and a half to three times the training throughput of an H100; the B300, its Blackwell Ultra refresh, raises memory to 288 GB and rents for about a dollar an hour more. In September 2026 the B200 rents for a median of $6.49 per GPU-hour on demand and the B300 for about $7.99, according to GetDeploying, with specialist clouds at $6 to $7 and the hyperscalers at $14 to $16 for the same silicon. Multi-year reserved commitments have been quoted as low as $2.25 per B200-hour. The catch is supply: only 16 percent of B200 listings and 11 percent of B300 listings are confirmed in stock, per AIMultiple, so a rate card is not an offer until the provider can name a delivery date.
Prices checked September 2026. Sources: GetDeploying, Thunder Compute, Spheron, IntuitionLabs. Free to cite with a link to this page.
Once you have asked a provider for Blackwell pricing directly, it will only work with you on its own terms. Send it to us first and you'll hear back within 24 hours from the person who will run your search, with every provider that can actually deliver quoted at once and benchmarked against the figures on this page. The provider you choose pays us, so it costs you nothing.
B200 and B300 Rental Prices in 2026
| Figure | B200 | B300 | Source |
|---|---|---|---|
| On-demand median | $6.49 per GPU-hour | $7.99 per GPU-hour | GetDeploying, B300 page, Sept 28, 2026 |
| Specialist cloud range, on demand | $5.99 to $6.82 | $7.10 to about $8 | Thunder Compute, B300 guide, Sept 2026 |
| Hyperscaler on demand | About $14.24 (AWS) to $16.11 (Google Cloud) | Up to about $17.80 | Thunder Compute, IntuitionLabs |
| Cheapest verified in stock | $3.75 | About $7.10 | GetDeploying, Thunder Compute |
| Spot | About $4.40; rose 48 percent between February and April 2026 | From about $4 per hour on the cheapest listings | Spheron, ColoPrice, GetDeploying |
| Reserved | 36-month commitments quoted as low as $2.25 | Median $5.71; 48-month commitments as low as $3.13 | IntuitionLabs, AIMultiple, Tech Insider |
| One GPU, one month, on demand | About $4,700 | About $5,800 | MCA calculation at the medians |
| One eight-GPU server, one month | About $38,000 | About $47,000 | MCA calculation at the medians |
| Buying instead | Eight-GPU HGX B200 server about $390,000 to $500,000 | About $50,000 to $60,000 per GPU; eight-GPU HGX B300 about $550,000 to $650,000 | GPU Smith, Thunder Compute, Tech Insider |
B200 vs B300: Which One Should You Rent?
The B300 is the B200 with more memory and more power on the same eight-GPU HGX footprint. Bandwidth is unchanged, so it helps when the model does not fit, not when the model is slow.
| Specification | B200 | B300 |
|---|---|---|
| Architecture | Blackwell; first HGX systems shipped late 2024 | Blackwell Ultra; broad availability from mid-2026 |
| Memory | 180 GB HBM3e (some vendors quote 192 GB) | 288 GB HBM3e |
| Memory bandwidth | About 8 TB/s | About 8 TB/s, unchanged |
| GPU power | Up to 1,000 W | Up to 1,400 W |
| Form factor | SXM only, eight per HGX B200 server, NVLink at 1.8 TB/s per GPU | SXM only, eight per HGX B300 server |
| Server power | Roughly 14 kW per eight-GPU server | Roughly 16 kW and up per eight-GPU server |
| Rental premium | Baseline | About $1 to $1.50 more per GPU-hour at providers listing both |
Sources: Thunder Compute, Runpod, Hyperstack. Server power figures are MCA estimates from GPU power ratings plus CPUs, memory, networking and fans.
Rent the B200 when
The model and its working set fit in 180 GB per GPU, or you are training across many GPUs anyway and memory per card is not the limit. It is the better price for most training and for inference on models up to roughly 100 billion parameters at FP4.
Rent the B300 when
A single-GPU or single-node deployment is memory-bound: the extra 108 GB lets one B300 serve a model, or a longer context, that would need two B200s. If the bottleneck is bandwidth rather than memory, the B300 will not help.
Stay on the H200 when
The workload fits in 141 GB and is inference rather than training. At $4.40 an hour the H200 is often the cheapest way to serve a large model. See H200 rental.
Stay on the H100 when
Cost per hour matters more than time to result, and the toolchain is mature. At half the price, the H100 still wins most fine-tuning and sub-80 GB inference. See H100 rental.
B200 vs H100 vs H200: Cost per Result, Not per Hour
| GPU | On-demand median, Sept 2026 | Memory | Training work per GPU vs H100 | Rough cost per unit of training |
|---|---|---|---|---|
| H100 | $3.25 | 80 GB | 1x | $3.25 |
| H200 | $4.40 | 141 GB | About 1.1 to 1.4x on training; more on memory-bound inference | About $3.15 to $4.00 |
| B200 | $6.49 | 180 GB | About 2.5 to 3x | About $2.15 to $2.60 |
| B300 | $7.99 | 288 GB | About 2.5 to 3x, higher where memory was the constraint | About $2.65 to $3.20 |
Cost per unit of training is an MCA calculation: the on-demand median divided by throughput relative to the H100, using the ranges published by ColoPrice and CloudZero. Real throughput depends on the model, precision and fabric.
Getting Allocation: Why Availability Matters More Than the Rate
Fewer than one Blackwell listing in five is confirmed in stock. A published $6 rate at a provider with no capacity is worth less than a $7 rate with servers energized and a delivery date in writing.
Ask for the delivery date in writing
"Available now" should mean a named quantity, a location and a date, with a remedy if the date is missed. "Subject to availability" is a quote, not capacity.
Quote several providers at once
Blackwell allocation is uneven: one specialist cloud has B300s energized while another is still waiting on servers. Only a search across providers finds who can deliver this quarter.
Consider monthly terms on a single server
Eight B200s on a monthly commitment, about $38,000 at the median, is now a standard offer, and it is often the fastest route to Blackwell capacity while a larger reservation is arranged.
Know the neoclouds
The specialist GPU clouds are where Blackwell capacity is cheapest and where allocation moves fastest. Who they are and how they differ is on neocloud providers.
Mind the term
A 36-month B200 reservation at $2.25 is a strong price, but it runs into the Rubin generation. Negotiate an upgrade path or a shorter term with a renewal option.
Verify before you commit
We confirm capacity with the provider before a client signs anything, because the trackers report listings, not servers.
What the Hourly Rate Leaves Out
| Item | What to check on Blackwell |
|---|---|
| Networking | Multi-node training needs 400 Gbps InfiniBand or equivalent per GPU; Blackwell moves enough data that a slow fabric wastes the premium you paid for the chip |
| Storage | Per-GB-month pricing and throughput; faster GPUs need faster storage to stay busy |
| Egress | Free at most specialist clouds, metered at the hyperscalers; checkpoints from Blackwell training runs are large |
| Billing increments and minimums | Per-second, per-minute or per-hour; minimum GPU counts and minimum runtimes |
| Software | FP4 and the second-generation Transformer Engine need current libraries; confirm the provider's image supports them or you lose the throughput advantage |
Rent or Buy a B200 Server?
An eight-GPU HGX B200 server costs about $390,000 to $500,000 to buy. At the on-demand median it rents for about $38,000 a month, so the crossover with buying comes at roughly a year of full use, before hosting. The complication is where to put it: a B200 server draws about 14 kW and a B300 server more, which is beyond most facilities built for 5 to 15 kW racks.
MCA calculation from the figures on this page; excludes power consumption, staff, networking and the residual value of the hardware. Free to cite with a link to this page.
Rent when
The project is under two years, demand is uncertain, or you cannot host a 14 kW-plus server. On Blackwell, renting also lets you move to Rubin without owning a depreciating asset.
Own and colocate when
The server will run for three years at high utilization and you can secure a facility for 40 kW-plus air-cooled racks. The AI and GPU colocation guide covers which facilities can, and we place those deployments too.
The long reservation
At the lowest quoted 36-month rates, a reserved B200 server is cheaper than owning one, without the capital or the facility. It is the best price on this page, and the one with the most to negotiate: delivery date, upgrade path and exit.
How We Get You Blackwell Capacity
You tell us the requirement: B200 or B300, GPU count, term, start date, storage and networking. We take it to every specialist GPU cloud and bare metal provider that can deliver Blackwell on your timeline, at the same time, so they compete; we confirm the capacity exists before you commit; we benchmark the quotes against the figures on this page; and we negotiate the rate, the delivery date, the term and the exit before you sign. A single server on a monthly term is welcome. The provider you choose pays us from the channel budget it would otherwise spend on its own sales team, so the service costs you nothing and the rate is not marked up. How it works.
Frequently Asked Questions
How much does it cost to rent a B200?
In September 2026 the median on-demand rate is $6.49 per GPU-hour across 21 providers, according to GetDeploying, with specialist clouds at $5.99 to $6.82 and the hyperscalers at $14 to $16. The cheapest verified in-stock rate is $3.75, spot runs about $4.40, and 36-month reserved commitments have been quoted as low as $2.25. One B200 costs about $4,700 a month at the median; an eight-GPU server about $38,000.
How much does it cost to rent a B300?
About $7.99 per GPU-hour at the median on demand, with specialist clouds from $7.10 and hyperscalers up to about $17.80. The reserved median is $5.71, and 48-month commitments have been quoted as low as $3.13. One B300 costs about $5,800 a month at the median; an eight-GPU server about $47,000.
What is the difference between the B200 and the B300?
The B300 is the Blackwell Ultra refresh of the B200: 288 GB of memory against 180 GB, 1,400 W against 1,000 W, the same 8 TB/s bandwidth, and about a dollar to a dollar and a half more per GPU-hour. It helps when a model or context does not fit on a B200; it does not help when the bottleneck is bandwidth.
Is the B200 worth it over the H100?
For training, usually yes: a B200 costs about twice an H100 per hour but does roughly two and a half to three times the work, so it is cheaper per result at the medians, and reserved discounts widen the gap. For fine-tuning and inference that fits in 80 GB, the H100 at half the hourly price is often still cheaper per job.
Should I rent a B200 or an H200?
For inference on a model that fits in 141 GB, the H200 at $4.40 is usually the cheapest option. For training, or inference that needs FP4 throughput or more than 141 GB, the B200 wins. The B200's FP4 precision roughly doubles inference throughput over FP8, which matters at scale.
Why are B200 prices so different between providers?
The same GPU runs about $6 on specialist GPU clouds and $14 to $16 on the hyperscalers, which bundle networking, storage and their own margins. Spot and small providers go lower still. Supply is the other factor: with only 16 percent of B200 listings confirmed in stock, providers with capacity charge for it.
Can I rent a single B200 server on a monthly term?
Yes. Eight B200s on a monthly commitment is a standard offer from specialist clouds and bare metal providers, at roughly $38,000 a month at the median, and it is often the fastest route to Blackwell capacity. We place single-server monthly deployments as well as reserved clusters.
How much power does a B200 server need?
An eight-GPU HGX B200 server draws roughly 14 kW, and an HGX B300 server 16 kW and up, from 1,000 W and 1,400 W per GPU respectively plus CPUs, memory, networking and fans. That is air-coolable at two or three servers per rack, but beyond what most facilities built for 5 to 15 kW racks can serve.
Is it cheaper to rent or buy a B200 server?
Over three years at full utilization, owning an eight-GPU B200 server and hosting it costs about $550,000, against about $890,000 to rent reserved at typical discounts and $1.36 million on demand. But the lowest quoted 36-month reserved rates come in near $470,000, cheaper than owning, without the capital or the facility.
How hard is it to get B200 or B300 capacity?
Harder than the rate cards suggest. Only 16 percent of B200 and 11 percent of B300 listings are confirmed in stock, according to AIMultiple, and allocation varies by provider and month. Quoting several providers at once, and getting the delivery date in writing, is how deployments actually land.
Will B200 prices fall when Rubin arrives?
Probably, following the pattern of every previous generation, though the timing depends on how fast Rubin ships and how strong demand stays. B200 on-demand rates rose about 16 percent in the year to September 2026 because supply lagged demand. A long reservation signed today should carry an upgrade path or an exit for that reason.
What does it cost to get Blackwell quotes through Metro Colo Advisory?
Nothing. The provider you choose pays us from its channel budget, the same way it pays its own sales team, and the rate is not marked up. If the best answer is a different GPU or the provider you already use, we say so.
Get B200 and B300 Quotes Through Us
Tell us B200 or B300, how many, the term, the start date, and the storage and networking you need. You'll hear back within 24 hours from the person who will run your search, with every provider that can actually deliver quoted at once and benchmarked against the figures on this page. If the better answer is an H200, an H100, or the provider you already use, we say so. No cost, no obligation, and the provider you choose pays us.
That is what keeps every provider competing for it, and on Blackwell it is how you find out who can deliver.
GPU rental: the GPU rental hub, GPU cluster rental, H100 rental, H200 rental, GB200 and GB300 NVL72 rental, monthly GPU rental prices, neocloud providers, CoreWeave competitors and bare metal.
What comes next: NVIDIA Rubin and the GPU roadmap, and what it means for Blackwell terms signed today.
Owning and hosting: AI and GPU colocation, direct to chip cooling and data center cost.