Tag: gradient accumulation
-

Supervised Fine-Tuning End to End: The First Real Run
A full supervised fine-tune twice over: a bare PyTorch loop and the same job through a trainer. Hyperparameters that work, plus how…

A full supervised fine-tune twice over: a bare PyTorch loop and the same job through a trainer. Hyperparameters that work, plus how…