الباحثون

Danni Yu

المنشورات 2

نسخة أولية وصول مفتوح

Forking: Sudden Overfitting Under Replay

Shanbin Yu, Shaoyang Guo, Haoran Zhao وآخرون · 2026

This paper studies forking, a generalization failure discovered in NanoGPT autoresearch. Under data replay, models with an over-encoding n-gram memory branch show a sharp separation of training and validation loss at epoch boundaries, resembling the shape of forks. We study this phenomenon in a controlled vanilla NanoG …

المؤلفون المشاركون