arXiv AI By Zefan Wang, Lincheng Li, Tianyu Yu, Yuan Yao

DRIFT: Refining Instruction Data via On-Policy Data Attribution

Read the original on arXiv AI →

arXiv:2606. 18307v1 Announce Type: cross Abstract: Optimizing the training data distribution for Supervised Fine-Tuning (SFT) dictates the capability of Large Language Models (LLMs).

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.