Tag: supervised fine-tuning
-

Supervised Fine-Tuning End to End: The First Real Run
A full supervised fine-tune twice over: a bare PyTorch loop and the same job through a trainer. Hyperparameters that work, plus how…
-

Masking in LLM Training: Loss, Causal and Padding
Three things share the name masking. Loss masking with -100, the causal mask the architecture applies, and the padding mask you build.…
-

LLM Fine-Tuning Explained: What Actually Changes Inside the Model
LLM fine-tuning runs the pretraining objective on your data with a loss mask. What actually changes, when to do it, and how…