Definition
Plain language
A high-end NVIDIA data-center chip used to train and run large AI models, a step up from the H100.
As stated in the literature
NVIDIA's Hopper-architecture data-center GPU with expanded high-bandwidth memory relative to the H100, used in large-scale training and serving.
Why it matters: Its expanded memory lets bigger models be trained and served more efficiently, directly affecting what's practical to build.
For example, a lab might train a large language model on a cluster of H200 chips because each one holds more high-speed memory than the previous generation.
Heard on the show
“Training the smaller model took five days of wall-clock on a hundred and ninety-two H200 GPUs.”Episode 080 — How a Two-Agent Trick Unlocked Large-Scale Training for Computer-Use Agents