Post-training
The fine-tuning steps (SFT, preference tuning) applied after pretraining to turn a base model into a helpful assistant.
The fine-tuning steps (SFT, preference tuning) applied after pretraining to turn a base model into a helpful assistant. (M01)