Glossary · Term

H200

← all terms

Definition

Plain language

A high-end NVIDIA data-center chip used to train and run large AI models, a step up from the H100.

As stated in the literature

NVIDIA's Hopper-architecture data-center GPU with expanded high-bandwidth memory relative to the H100, used in large-scale training and serving.

Why it matters: Its expanded memory lets bigger models be trained and served more efficiently, directly affecting what's practical to build.

For example, a lab might train a large language model on a cluster of H200 chips because each one holds more high-speed memory than the previous generation.

Heard on the show

“Training the smaller model took five days of wall-clock on a hundred and ninety-two H200 GPUs.”
Episode 080 — How a Two-Agent Trick Unlocked Large-Scale Training for Computer-Use Agents

Related terms