الباحثون

Zhi-Qin John Xu

المنشورات 3

نسخة أولية وصول مفتوح

Probability-Signature Dynamics: Unpacking Modular Addition Learning Within Two-Layer Networks

Yunji Wang, Junjie Yao, Linyu Liu وآخرون · 2026

Neural networks trained on modular addition tasks often develop Fourier-structured representations that support exact generalization. While prior work has identified these Fourier circuits, the mechanism by which gradient-based training selects them from the data distribution remains unclear. We address this question u …

نسخة أولية وصول مفتوح

Weight Decay and Neuron Condensation: A Three-Stage Analysis of Two-Layer ReLU Networks

Cheng Xu, Pengxiao Lin, Zhangchen Zhou وآخرون · 2026

Weight decay is widely used as a regularization technique in neural network training, yet its role in neuron condensation (parameter direction alignment) remains unclear. Starting from a parameter initialization in the neural tangent kernel regime, we characterize training dynamics under weight decay through three stag …

المؤلفون المشاركون