الباحثون

Yanzhe Chen

المنشورات 3

نسخة أولية وصول مفتوح

Learn the Directions, Normalize the Gains: Post-Training Normalization for LoRA

Zailong Tian, Yanzhe Chen, Zhuoheng Han وآخرون · 2026

While Low-Rank Adaptation (LoRA) enables efficient task specialization, its learned updates can compromise capabilities beyond the target task. We identify \textbf{adaptation imbalance}: a few singular directions dominate the trained update, leaving its performance sensitive to how gains are allocated. We argue that \t …

نسخة أولية وصول مفتوح

Muon Sublates the Edge of Stability in LLM Pretraining

Yanzhe Chen, Qifang Zhao, Xiaoxiao Xu وآخرون · 2026

Muon is increasingly used for language-model pretraining, yet its large-step dynamics are not captured by the classical edge-of-stability (EoS) picture of gradient descent (GD). In GD, loss neutrality, equal-magnitude update reversal, and marginal stability meet at a single learning-rate-dependent edge. We show that Mu …

نسخة أولية وصول مفتوح

PaperDoctor: Evidence-Grounded and Actionable Feedback for Scientific Papers in Progress

Kevin Qinghong Lin, Siyuan Hu, Pan Lu وآخرون · 2026

Autoresearch agents are reshaping the research ecosystem, but they can also let flawed claims enter the literature at scale. Human advisors catch such issues in drafts through careful, traceable feedback, yet advisor-style assessment requires extensive manual effort and does not scale. To shift automated paper assessme …

المؤلفون المشاركون