Post-training

The smaller training phase that shapes a model’s behaviour rather than its knowledge.

Also known as alignment

Supervised fine-tuning on example conversations teaches the assistant format; reinforcement learning from human feedback then tunes toward responses people rate highly.

Post-training is a thin layer over an enormous pretrained base, and it is the layer vendors compete on hardest. It explains why models with similar underlying capability can have very different personalities and refusal behaviour.