Glossary · Term

H100

← all terms

Definition

Plain language

A high-end NVIDIA data-center GPU used to train and serve large AI models.

As stated in the literature

NVIDIA's Hopper-architecture data-center GPU, widely used in frontier-model training and serving stacks.

Why it matters: It's the hardware nearly all modern frontier training is done on, and its availability and price largely determine who can build state-of-the-art models.

For example, training a 70B model from scratch typically requires a cluster of thousands of H100s running for weeks.

Heard on the show

“Llama-three-point-one-eight-B on an H100 — the most standard, most commodity LLM serving deployment in the world.”
Episode 027 — When AI Agents Build the Serving Stack: A Bet on Bespoke Infrastructure

Related terms