Definition
Plain language
A high-end NVIDIA data-center GPU used to train and serve large AI models.
As stated in the literature
NVIDIA's Hopper-architecture data-center GPU, widely used in frontier-model training and serving stacks.
Why it matters: It's the hardware nearly all modern frontier training is done on, and its availability and price largely determine who can build state-of-the-art models.
For example, training a 70B model from scratch typically requires a cluster of thousands of H100s running for weeks.
Heard on the show
“Llama-three-point-one-eight-B on an H100 — the most standard, most commodity LLM serving deployment in the world.”Episode 027 — When AI Agents Build the Serving Stack: A Bet on Bespoke Infrastructure