H200 GPU Rental Prices & Live Availability
As of , no single H200 has a current in-stock signal. The lowest listed price is $2.00 per GPU-hour on PrimeIntellect, but that provider does not expose a current stock signal. GPU Finder compares 9 providers for this size and refreshes availability hourly. Methodology.
Structured offer basis: 36 current 1× H200 listings range from $2.00 to $6.46 USD per physical GPU-hour. Whole-node prices for other GPU counts are not mixed into this range.
Cheapest single H200 rental
verified- Is now a good time to rent an H200 cloud GPU?
- ▲ near its 6-month high — The 8× H200 on-demand floor is $15.97/hr ($2.00/GPU-hr) — 6-month range $3.95–$15.97/hr for 8× nodes. In short, prices are near their 6-month high, so it may pay to wait or lock in a spot rate.
- What is the cheapest H200 cloud GPU available right now?
- No single H200 is confirmed in stock across tracked providers at the moment.
- Which cloud providers most reliably have the H200 in stock?
- Most reliably in stock over the last 30 days: PrimeIntellect, AceCloud. Frequently waitlisted: Vast, Hyperstack.
Prices are the cheapest listed on-demand rates; “in stock” means the provider reported capacity on our last hourly check. How we compute this →
H200 rental prices by provider
- datacrunch__1xH2…41GB__44__182Stale · 3d ago100% · 30d$2.00/hr1× H20044 vCPU182 GBFree egressFI, IS
Get notified when H200 drops in price or reliability changes
One email per change, max once a day. We send a confirmation link first; one-click unsubscribe in every email.
H200 (1× GPU) availability — last 7 days
1 of 6 providers kept 1× H200 ≥80% available all week. Polled hourly · 6 providers tracked · 7-day window.
| Provider | Tue29 | Wed30 | Thu1 | Fri2 | Sat3 | Sun4 | Mon5 | Week |
|---|---|---|---|---|---|---|---|---|
| PrimeIntellect | 100% | |||||||
| AceCloud | 26% | |||||||
| Nebius | 16% | |||||||
| Lium | 7% | |||||||
| Runpod | 0% | |||||||
| Vast | 0% |
30-day reliability depth
GPU Finder does not treat availability as a static yes/no flag. For H200, current stock badges are paired with a 30-day reliability score based on our hourly stock checks — the share of tracked time each listing was reported available. How we compute this →
- 105
- listings with a 30-day score
- 72%
- of listings on this page have a score
- 30d
- longest stock history on this page
Strongest reliability signals here: PrimeIntellect (100%, 30d covered), AceCloud (84%, 30d covered), Lium (68%, 30d covered).
Caveat: scores stay hidden until a listing has at least 48 tracked hours, and some provider APIs expose coarse capacity levels instead of exact stock counts. Use the score as historical depth beside current availability, not as a guarantee that a GPU will still be allocatable when you click through.
Price History & Comparison
Full H200 price history (21 months) →Spot pricing decision guide
Treat spot as a risk-adjusted capacity decision, not just a cheaper number.
per GPU-hour
63% below on-demand floor
daily median for the same offers with complete history
median $2.50/GPU-hr
Cheapest trusted spot / interruptible capacity is on Vast; 0 providers currently report spot or interruptible stock.
Nebius has kept spot or interruptible capacity in stock most reliably over the last 30 days (18%). See the availability section for the broader stock history.
The discount is visible, but availability, reliability, or volatility argues for more caution.
Good fit: checkpointed batch jobs, flexible training, and experiments that can restart. Avoid for production serving, deadline-bound runs, or jobs that cannot tolerate eviction.
Source caveat: Runpod exposes explicit spot fields; Vast is labeled interruptible/bid-floor. Lambda is treated as on-demand-only unless a verified spot field is added.
H200 spot and interruptible pricing
H200 spot and interruptible GPU pricing can vary by provider, stock, and interruption risk. GPU Finder shows verified Runpod spot rows and Vast interruptible/bid-related rows beside on-demand pricing where current source data supports it.
Vast interruptible
$0.74/GPU-hr
vs on-demand $4.00/GPU-hr · 81% below on-demand
Vast interruptible H200 rowsGood fit: checkpointed batch jobs, flexible training, and experiments that can restart. Avoid for production serving, deadline-bound runs, or jobs that cannot tolerate eviction. Compare Runpod vs Vast spot and interruptible before renting.
H200 spot and interruptible FAQ
Who has H200 spot GPU pricing?
GPU Finder shows verified Vast interruptible/bid-related H200 rows where current source data supports it. Check the pricing table for on-demand and spot or interruptible offers from other providers.
What is the difference between H200 spot and interruptible pricing?
Runpod exposes explicit spot fields tied to Secure Cloud and Community Cloud. Vast uses a marketplace model, so its discounted H200 capacity is labeled interruptible or bid-related rather than generic spot.
Should I choose the cheapest H200 interruptible offer?
Only if the workload can tolerate interruption. H200 spot and interruptible prices are best for checkpointed batch jobs, flexible training, and experiments. Avoid them for production serving, deadline-bound runs, or non-checkpointed workloads.
About the H200
The NVIDIA H200 is a memory-upgraded variant of the H100, sharing the same Hopper compute architecture but with 141 GB of HBM3e at 4.8 TB/s. This 76% increase in memory capacity over the H100 directly translates to higher throughput on memory-bound inference. For teams already on Hopper but bottlenecked by 80 GB, the H200 is a drop-in upgrade with no software changes.
Key Specifications
| Architecture | Hopper (GH100) |
| GPU Memory | 141 GB HBM3e |
| Memory Bandwidth | 4.8 TB/s |
| FP16 Tensor Core | 1,979 TFLOPS |
| TDP | 700W (SXM) |
| Interconnect | NVLink 4.0 (900 GB/s) |
| Release Year | 2024 |
Cloud Pricing Context
H200 on-demand pricing starts around $2.14/hr and runs up to $3.59/hr. Pricing is settling between H100 and B200 rates. The H200 offers strong value for inference where extra memory eliminates tensor parallelism across multiple H100s, effectively halving infrastructure costs.
Best For
- Inference serving of 70B+ parameter models on a single GPU
- Memory-bound workloads where HBM3e bandwidth is the limiting factor
- Teams upgrading from H100 who need more VRAM without requalifying software
- KV-cache-heavy serving (long-context LLMs) where 141 GB keeps throughput high
8× H200 server pricing and bundle math
Live-derived multi-GPU estimates for 8× H200 (1128 GB aggregate VRAM).
Normalized from live on-demand rows by GPU count.
Cheapest live 8-GPU instance.
Approx. 24 × 30 at the bundle hourly rate.
GPU Finder currently tracks a live 8-GPU H200 instance from $15.97/hr ($11498.40/mo at 24×30). Use the table below to compare exact configurations and availability.
Networking, CPU, host memory, storage, and egress vary by provider and are not folded into this estimate. Prices are for on-demand cloud instances unless the provider table shows a spot or interruptible rate.
8× H200 server pricing FAQ
How much does an 8× H200 server cost per hour?
Based on current live rates, the cheapest 8× H200 bundle starts around $15.97/hr, which is roughly $11498.40/mo at 24×30 utilization. The shorthand "8xh200 price" captures the most common search form. Per-GPU H200 rates at $2.00/GPU-hr anchor the estimate, with exact 8-GPU instances varying by provider.
What is the monthly price for an 8× H200 server?
An 8× H200 server at $15.97/hr works out to roughly $11498.40/mo at full 24×30 utilization. Egress, storage, and committed-use discounts are extra.
How does 8× H200 pricing compare to 8× B200?
H200 uses the same Hopper architecture as H100 with 141 GB HBM3e per GPU. 8× H200 is often easier to find than 8× B200 today and can be cheaper on a per-hour basis, though B200 offers more future throughput per GPU.
Where can I rent a single H200 GPU?
Use the live H200 provider table on this page to compare single-GPU and multi-GPU H200 listings. Availability is expanding through 2026 as production ramps.
Is 8× H200 cheaper than 8× H100?
H200 usually costs more per GPU than H100, but the extra 61 GB of memory per GPU can reduce the total GPU count needed for memory-bound inference, so the effective 8-GPU job cost depends on workload fit.
H200 by node size
Dated availability research
Cite the GPU Cloud Availability Index for frozen H200 provider reliability rankings, coverage hours and downloadable data.