Supervised fine-tuning (SFT)
Training on curated prompt-response pairs so the model imitates demonstrated behavior — the first stage of turning a base model into an assistant, and the standard tool for teaching format and style.
Training on curated prompt-response pairs so the model imitates demonstrated behavior — the first stage of turning a base model into an assistant, and the standard tool for teaching format and style.