Tools / H100 Price Tracker
H100 & H200 price tracker
The cheapest rentable NVIDIA H100 and H200 cloud prices right now, compared live across Vast.ai, Verda, and Lyceum. Single-GPU instances only, spot and on-demand, normalized to $/GPU/hr. Prices refresh automatically every 30 seconds while the page is open.
Collecting history for bid guidance…
| GPU | Provider | Type | $/GPU/hr |
|---|---|---|---|
| Fetching live prices… | |||
Spot is interruptible capacity; Vast.ai spot prices are live minimum bids. Excludes storage and egress.
What this tracks
The board lists the cheapest currently-rentable single-GPU H100 and H200 instances on three clouds with fully public pricing, sorted by price per GPU-hour. It deliberately tracks 1× instances because that is what most fine-tuning and inference experiments actually rent; multi-GPU cluster pricing follows different supply and is not comparable per GPU. Spot rows are interruptible capacity — the provider can reclaim the machine — while on-demand rows are dedicated instances that run until you stop them.
Where the prices come from
Vast.ai prices are live marketplace offers from its public search API, filtered to verified, rentable hosts; the spot price shown is the current minimum bid on an interruptible offer, which is the price you would pay to take the machine right now. Verda (formerly DataCrunch) prices come from its public API, including its true spot tier and dynamic on-demand pricing. Lyceum publishes flat list prices and has no spot tier, so its rows are always on-demand. All prices exclude storage and egress. We are not affiliated with any of the three providers; refresh happens through our server every 30 seconds, and nothing about you is sent anywhere.
Spot vs on-demand: which should you rent?
Spot capacity is 40 to 80% cheaper and perfectly good for anything that can survive an interruption: fine-tuning with regular checkpoints, batch inference, evaluations, and development. On-demand is for work that must not die mid-run — production endpoints, long uninterruptible training, demos. A common pattern is to develop and train on spot, then serve on the cheapest reliable on-demand instance. If you are estimating how much GPU memory your job needs first, our LLM VRAM calculator and fine-tuning memory calculator answer that.
Why the same GPU costs 10× more elsewhere
The chips are identical; the business models are not. Marketplace hosts compete each other's prices down and sell interruptible capacity for whatever the market bears, specialist GPU clouds price for reliability and support, and hyperscalers price H100s at $6–12/hr for enterprise integration and compliance. Form factor matters too: SXM parts with NVLink command a premium over PCIe cards, and the H200's extra memory (141 GB vs 80 GB) adds roughly a dollar per hour. If a price here looks too good, remember what spot means — a $0.60 H100 is real, but it is a minimum bid on a machine someone can outbid you for.
Bid guidance: how much should you bid?
On marketplaces like Vast.ai, the spot price is a minimum bid: you name a price, you hold the machine while your bid stays above the market, and you lose it when someone outbids you. So the practical question is not "what is the cheapest price" but "what should I bid so the cheapest machine stays mine". The chart answers it from history: the blue line traces the cheapest observed price — never an average, never other offers — drawn from every roughly 30-second market sample on the 1H, 6H, and 24H views. The 7-day view aggregates to hourly points but shades each hour's full low-to-high range, so no price movement is hidden at any zoom. The dashed guides mark the bid levels that would have beaten the cheapest price in 50%, 75%, and 90% of those raw samples; they are computed from the samples, not the drawn line, so zooming never changes the guidance. Bid at the aggressive (50%) level and you win half the time but get interrupted often; bid at the 90% level and you pay a few cents more per hour for a bid that almost always held. The suggested bid above the chart is the 90% level, recomputed live for whatever range and filters you have selected. History is retained for 30 days.
How much does an H100 cost per hour in 2026?
The live board above is the source of truth, but the ranges are fairly stable. These are the single-GPU prices we observe across the three tracked providers as of July 2026, against typical hyperscaler list prices for the same chip:
| GPU | Spot / interruptible | On-demand | Hyperscaler on-demand |
|---|---|---|---|
| H100 80GB | $0.60 – $1.60/hr | $1.75 – $3.60/hr | $6 – $12/hr |
| H200 141GB | $1.00 – $1.40/hr | $2.90 – $4.00/hr | $7 – $14/hr |
In short: renting an H100 costs from about $0.60/hr on interruptible spot capacity, $2–4/hr on-demand at specialist GPU clouds, and $6–12/hr at the big three clouds. An H200 adds roughly $0.40–1/hr over the H100 at the same tier for 76% more memory.
Frequently asked questions
How much does an H100 cost per hour?
It depends heavily on where you rent it. On GPU marketplaces, interruptible spot capacity for a single H100 can dip below $1/hr, dedicated on-demand instances on specialist clouds typically run $2 to $4/hr, and hyperscalers charge $6 to $12/hr for the same chip. This tracker shows the live low end of that range across three providers with public pricing.
What is a spot GPU instance?
Spot (or interruptible) instances are spare capacity sold at a steep discount with the catch that the provider can reclaim them at short notice. They are 40 to 80% cheaper than on-demand and suit fault-tolerant work like checkpointed fine-tuning or batch inference, but not production endpoints that must stay up.
Why do H100 prices differ so much between providers?
The hardware is identical, so price differences come from the business model: marketplace hosts compete on price and sell interruptible capacity cheaply, specialist GPU clouds price for reliability and support, and hyperscalers price for enterprise integration. Interconnect (SXM vs PCIe), region, and demand also move prices hour to hour.
How much does an H200 cost compared to an H100?
Expect to pay roughly $0.40 to $1 more per hour for an H200 at the same provider and tier. The H200 is the same Hopper GPU with 141 GB of HBM3e instead of 80 GB, so the premium buys memory headroom: larger models, longer context, and bigger batches without sharding across two cards, which can make one H200 cheaper than two H100s.
Is it cheaper to rent or buy an H100?
An H100 card costs roughly $25,000 to $30,000 before the server, power, and networking around it. At a rented rate of $2/hr you could run it for about 15,000 hours — around two years of continuous use — before matching the purchase price alone. Renting wins unless you can keep the card busy around the clock for years and are equipped to host it.