arXiv Machine Learning By Mark Rofin, Aditya Varre, Nicolas Flammarion

(How) Learning Rates Regulate Catastrophic Overtraining

Read the original on arXiv Machine Learning →

arXiv:2604. 13627v2 Announce Type: replace Abstract: Supervised fine-tuning (SFT) is a common first stage of LLM post-training, teaching the model to follow instructions and shaping its behavior as a helpful assistant.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv Machine Learning.