arXiv AI By Yunze Han

A Systematic Evaluation of Trajectory Data Curation for LoRA Fine-Tuning of Code Agents

Read the original on arXiv AI →

arXiv:2607. 17205v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) of open-weight LLMs on expert agent trajectories has emerged as a prominent approach to building capable code agents without reliance on proprietary models.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.