arxivcs.AIcs.SE2026-07-19
A Systematic Evaluation of Trajectory Data Curation for LoRA Fine-Tuning of Code Agents
Supervised fine-tuning (SFT) of open-weight LLMs on expert agent trajectories has emerged as a prominent approach to building capable code agents without reliance on proprietary models. A central yet underexplored question is how trajectory quality and quantity jointly shape mode…