NVIDIA

Cloud GPU · verified 2026-09-05

NVIDIA H100

Hopper80 GB

The default datacenter hour. 80 GB of HBM3 is enough for most 70B work if you manage the KV cache. It is not a bargain because the FLOPs slide is pretty. It is a bargain when the job already fits.

VRAM80 GB
From /GPU/hr$1.73
ArchitectureHopper
VendorNVIDIA

In brief

What to know before you rent

Catalog listings

H100 Pricing and Availability

Published on-demand rates first. Spot and waitlists are labeled. Confirm the live rate before you provision.

ProvidersRegionBilling typeInterconnect/GPU/hr
Vast.aiIn stockUS-EastOn-demandPCIe$1.73Open provider
Packet.aiLimitedUS-WestOn-demandNVLink$2.19Open provider
RunPodIn stockUS-EastOn-demandNVLink$2.39Open provider
LambdaIn stockUS-WestOn-demandNVLink$2.49Open provider
FluidstackIn stockEU-WestOn-demandPCIe$2.69Open provider
KoyebIn stockEU-WestOn-demandNVLink$2.85Open provider
HyperstackIn stockEU-WestOn-demandNVLink$2.99Open provider
CrusoeIn stockUS-WestOn-demandNVLink$2.99Open provider
NebiusIn stockEU-NorthOn-demandNVLink$3.10Open provider
Voltage ParkIn stockUS-WestOn-demandNVLink$3.35Open provider
Together AIIn stockUS-WestOn-demandNVLink$3.49Open provider
CoreWeaveLimitedUS-EastOn-demandInfiniBand NDR$3.89Open provider
AWSLimitedUS-EastOn-demandNVLink$4.10Open provider
Google CloudLimitedUS-WestOn-demandNVLink$4.85Open provider
AzureLimitedUS-EastOn-demandNVLink$4.98Open provider

Our reading

Who this GPU is for

I rent H100 when I can fill the SMs and the 80 GB envelope. I do not rent it as a status GPU. Marketplace H100s at meme prices are a different product from a Lambda or CoreWeave SXM node. Compare the SKU, not the three characters on the badge.

Rent it if Fine-tunes and serving that fit 80 GB, single-node training that actually uses NVLink, teams who can checkpoint and who will read SXM vs PCIe before they paste a card.

Skip it if Context windows that blow the KV cache, multi-node jobs that need a reserved fabric, or anyone shopping a 4090 workload with Hopper money.

How we compare hours

Datasheet

Technical specifications

NVIDIA H100 · Per GPU · SXM

Compute · dense

FP8
1,979 TFLOPS3,958 with sparsity
FP16 / BF16
989.5 TFLOPS1,979 with sparsity
TF32
494.5 TFLOPS989 with sparsity
INT8
1,979 TOPS3,958 with sparsity
FP32
67 TFLOPS
FP64
34 TFLOPS67 TFLOPS Tensor Core
Precision support
FP8, FP16, BF16, TF32, FP32, FP64, INT8

Memory

Capacity
80 GB HBM3
Bandwidth
3,350 GB/s
Bus width
5,120-bit
ECC
Yes

Silicon

Architecture
Hopper
Process
TSMC 4N
Transistors
80 billion
Shader cores
16,896CUDA cores
Matrix cores
5284th-gen Tensor cores
Compute units
132SMs

Fabric and host

GPU interconnect
NVLink 900 GB/s
Host interface
PCIe 5.0 x16128 GB/s

Power

Board power
700 WSXM, configurable

Platform

Partitioning
MIG, up to 7 instances10 GB each on SXM 80 GB

Table is H100 SXM, the SKU catalogs mean when they say “an H100.” H100 PCIe is ~350 W with slower HBM. H100 NVL is 94 GB and a 600 GB/s NVLink bridge, useful for 70B inference, not a silent upgrade. Confirm the form factor on the listing.

Trade-offs

Pros and cons of H100

What this GPU does well, and when the hour is the wrong buy.

Strengths

  • The SKU every catalog actually stocks. You can compare an hour, not a waitlist
  • FP8 Transformer Engine is the reason this generation still trains at a sane $/step
  • MIG exists when you want to split an 80 GB card instead of overpaying for a second one
  • NVLink 900 GB/s on SXM is the difference between 8 GPUs and 8 expensive heaters

Limits

  • 80 GB is the wall. Long context and fat Adam states do not care about your TFLOPS screenshot
  • PCIe H100 is slower memory and less interconnect, same marketing name
  • 700 W SXM needs a real node. A cheap host with a sad PSU is not a discount, it is a crash
  • Sparse peak numbers do not show up in your trainer unless the kernel path is sparse

Form factor

What's the difference between the H100 SXM, PCIe and NVL?

Same badge, different memory, bandwidth or interconnect. Confirm the SKU on the listing.

VariantMemoryMemory bandwidthInterconnect
H100 SXM80 GB3,350 GB/sNVLink 900 GB/s
H100 PCIe80 GB2,000 GB/sNVLink 600 GB/s
H100 NVL94 GB3,938 GB/sNVLink 600 GB/s

H100 SXM, PCIe and NVL are different products

H100 is three products wearing one badge. SXM is the HGX node: 80 GB HBM3, 3.35 TB/s, NVLink 900 GB/s, up to 700 W. PCIe is the card you can bolt into a random chassis: less bandwidth, less fabric, often a different memory generation. NVL is the 94 GB twin for 70B-class inference in power-capped halls.

If a marketplace row says H100 and the price looks like a consumer card, you are probably not buying the SXM row Lambda would sell you. Compare hours inside a SKU. Mixing them is how people save twenty cents and lose a training run.

80 GB is the H100 decision, not 1,979 TFLOPS

Most people landing here are fitting a 70B, serving a LoRA, or trying to keep a KV cache in HBM. Hopper’s SM count is the same family as H200. The upgrade is memory. If your batch already fits, you are paying H200 rent for headroom you will not touch.

Adam states, gradients, and long context are what blow 80 GB, not the marketing FLOPs. If the optimizer does not fit, extra TFLOPS are a very expensive fan. Buy memory first. Then fill the SMs.

H100 on-demand prices vs waitlisted Blackwell

On-demand H100 is the workhorse of this index: enough providers that you can walk away. That is the product. A waitlisted Blackwell at a fantasy FLOPs per watt is not a cheaper H100.

I still open Vast.ai for unverified floor prices and RunPod Secure when I want the job to finish. Lambda when the interconnect is the job. Hyperscalers when the data cannot leave. The GPU name stays H100. The product does not.

How I actually rent it

  1. 01

    Name the SKU

    SXM, PCIe, or NVL. GPU count. NVLink or not. If the listing will not say, leave.

  2. 02

    Fit the 80 GB

    Model + states + KV cache. If it spills, you want H200 or more GPUs, not a denser kernel fantasy.

  3. 03

    Compare on-demand only

    Spot and waitlists are labeled on this site. Do not mix them into the hour you budget.

FAQ

H100 questions

Who has the cheapest H100?+

As of September 5, 2026, the cheapest verified in-stock on-demand H100 tracked by GPUBeacon is $1.73 per GPU per hour from Vast.ai.

How much does the H100 cost per month?+

As of September 5, 2026, at 720 hours per month, one H100 costs an estimated $2,153 at the median on-demand price in this sample. Cheapest verified in stock: $1,246 per month on-demand.

Where can I rent an H100?+

GPUBeacon currently tracks H100 configs from 15 providers. The cheapest verified in-stock on-demand configs come from Vast.ai, RunPod, Lambda and Fluidstack. See the full price comparison above for every provider and config.

Are H100 prices going up or down?+

GPUBeacon publishes a catalog sample, not a 90-day price index. As of September 5, 2026, the median on-demand H100 in this sample is $2.99 per GPU per hour. The cheapest verified in-stock hour is $1.73.

How many H100 GPUs do I need?+

One H100 runs models up to roughly 120B parameters at 4-bit quantization or 31B at 16-bit, assuming a 32K context. Larger models run across multiple GPUs.

Why choose the H100?+

80 GB HBM3 with 900 GB/s NVLink for multi-node training. FP8 Transformer Engine for improved training throughput over A100. The standard choice for large-scale AI training and high-throughput inference.

When is the H100 not a good fit?+

May be more than needed for single-GPU inference on models under 30B parameters. Consider the L40S or A10 for cost-effective inference workloads.

What's the difference between the H100 SXM, PCIe and NVL?+

The H100 ships in 3 versions with different memory, bandwidth or interconnect. The figures at the top of this page are for the H100 SXM.

Is H100 still worth renting in 2026?+

Yes, when the job fits 80 GB and you need a SKU that is actually in stock. Blackwell is faster. Blackwell is also a waitlist with a power bill. H100 is the hour you can buy this week.

H100 or H200?+

H200 if the extra 61 GB of HBM3e shortens the job enough to beat the H100 hour, usually KV cache and fat states. If you already fill 80 GB with compute-bound work, stay on H100.

Do NVIDIA’s TFLOPS numbers include sparsity?+

The datasheet Tensor rows are with sparsity. Dense FP8 on H100 SXM is 1,979 TFLOPS. If your stack is not sparse, believe the dense column.

Can I MIG an H100 on a cloud?+

The silicon supports up to 7 instances. The provider may not expose it. If you need a slice, confirm the product page, do not assume a marketplace pod is MIG.

Keep comparing

Alternatives to NVIDIA H100

Same job, different envelope. Compare VRAM, bandwidth and the hour before you lock a badge.

NVIDIA B300

288 GB HBM3e · 2.4x the bandwidth of the H100

From$7.50/ GPU / hr

Compare

NVIDIA H200

141 GB HBM3e · 43% more bandwidth than the H100

From$3.00/ GPU / hr

Read the verdict

NVIDIA RTX 5090

32 GB GDDR7 · 47% less bandwidth than the H100

From$0.34/ GPU / hr

Compare

NVIDIA A100

80 GB HBM2e · 39% less bandwidth than the H100

From$1.19/ GPU / hr

Read the verdict

AMD MI300X

192 GB HBM3 · 58% more bandwidth than the H100

From$1.99/ GPU / hr

Compare

Browse all GPUs

Sources

Where these numbers come from

TFLOPS and TDP are from the vendor datasheet. Hourly prices are from the GPUBeacon catalog sample. Open the originals, then confirm the live SKU with the provider before you rent.