Definition
Plain language
The extra rounds of training that turn a finished base model into something useful and well-behaved, after the big initial training on raw text.
As stated in the literature
The stages applied after pretraining — supervised fine-tuning, RLHF, instruction tuning, and related methods — that adapt a base language model into an assistant or specialist; where alignment and most behavioral shaping happen.
Also called: post-trained, post-train
Why it matters: It is where most of a model's helpfulness, manners, and safety come from, so it largely determines how the finished assistant behaves.
For example, post-training is the stage that turns a raw text-predictor into a chatbot that politely answers questions and refuses harmful requests.
Heard on the show
“They take an open model, OLMo-3-32B-Think, and walk it through post-training stage by stage.”Episode 246 — 160 Perfect Refusals, And The Refusals Were The Leak