الباحثون

Xinyi Luo

المنشورات 1

نسخة أولية وصول مفتوح

Every Batch Is Its Own Validation Set: Leave-One-Out Gradient Matching for Online Data Selection in LLM Fine-Tuning

Hongyu Chen, Xinyi Luo, Ming Zhao وآخرون · 2026

Online batch selection fine-tunes a language model on the most useful part of each candidate batch. Selectors that match the gradient of the candidate batch are attractive because they need no held-out data, yet they rarely beat training on the whole batch. We show why. In-sample gradient matching uses every example as …

المؤلفون المشاركون