NVIDIA Hopper · Tensor Core GPU

NVIDIA H100 Tensor Core GPU

The proven Hopper-generation workhorse. The H100 SXM5 pairs 80GB of HBM3 with fourth-generation Tensor Cores and NVLink, and remains the most widely deployed accelerator for large-model training and high-throughput inference — with a deep secondary market that keeps lead times and pricing competitive.

  • Hopper
  • 80GB HBM3
  • 700W SXM5
  • Most available

H100 specifications

Key specifications for the NVIDIA H100 Tensor Core GPU. Hard specs mirror the live catalog; confirm exact revision and configuration at quote.

NVIDIA H100 Tensor Core GPU — reference specifications
SpecificationH100
ArchitectureHopper (GH100, TSMC 4N)
GPUs per unit1 (8 per HGX/DGX server)
Memory80 GB HBM3
Memory bandwidth3.4 TB/s
FP8 tensor3,958 TFLOPS (with sparsity)
FP16 / BF16 tensor1,979 TFLOPS (with sparsity)
NVLink900 GB/s (4th-gen NVLink)
Board power (TDP)700 W
Form factorSXM5
CoolingAir or direct-to-chip liquid
Typical lead time10–14 wk
Indicative priceFrom $26,500 / unit

What the H100 is for

The H100 is the default choice when you want a battle-tested accelerator with a broad software stack and predictable supply. Its 80GB of HBM3 comfortably holds most models up to the tens-of-billions of parameters range with tensor and pipeline parallelism, and FP8 Transformer Engine support roughly doubles throughput over A100-class hardware.

Foundation-model training

Scales across NVLink and 400G/800G fabrics into large training clusters. The proven choice for training runs where the software stack maturity matters.

High-throughput inference

80GB fits most production models; FP8 delivers strong tokens-per-second per dollar for served endpoints and batch inference.

Fine-tuning & research

Widely available new, used and refurbished — the pragmatic pick for teams standing up capacity on a budget and timeline.

Power & cooling for the H100

Density decides the facility. Size the electrical capacity, heat rejection and cooling method before you rack a single H100 — our free tools turn the specs above into facility numbers.

Source the H100

Add it to your quote list — free, no commitment — and we confirm a firm, itemized price, availability and lead time after review, usually within a business day.

NVIDIANVIDIA H100 Tensor Core GPU80GB HBM3 · 3.4TB/s · 700W · SXM510–14 wkFrom $26,500 / unit
View full product

H100, answered

How much memory does the NVIDIA H100 have?

The H100 SXM5 has 80GB of HBM3 with about 3.4 TB/s of memory bandwidth. A PCIe H100 NVL pairs two boards for 188GB combined; the newer H200 raises a single module to 141GB of faster HBM3e.

How much power does an H100 use?

The H100 SXM5 has a rated board power (TDP) of 700W. A full 8-GPU HGX/DGX H100 server draws roughly 10kW once CPUs, memory, NICs, fans and PSU losses are added — size cooling and power to that figure, not to raw GPU TDP.

What is the difference between the H100 and H200?

They share the same Hopper GH100 silicon and identical compute (FP8/FP16 TFLOPS). The H200 upgrades memory to 141GB HBM3e at 4.8 TB/s, versus 80GB HBM3 at 3.4 TB/s on the H100 — a large win for memory-bound inference and long-context workloads, at the same 700W.

How much does an H100 cost and what is the lead time?

Indicative pricing starts around From $26,500 / unit, with a typical lead time of 10–14 wk — confirmed per line at quote. A deep secondary market means used and refurbished units are often available faster and cheaper.

Turn the spec into a scoped quote

Our solutions engineers size compute, power and cooling together and confirm real lead times — no payment, no commitment, quotes back in about one business day.

Want a second opinion on a build?

Our engineers scope power, cooling and compute together. No payment, and quotes come back in about a business day.

Talk to an engineer