الباحثون

Junze He

المنشورات 1

نسخة أولية وصول مفتوح

Understanding the Weight Averaging Mechanism in LLM Training for Post-Training Quantization

Hanzhang Wang, Tianqi Shen, Zonglin Liu وآخرون · 2026

Large language models (LLMs) are typically pretrained in high precision but increasingly deployed with low-precision post-training quantization (PTQ). Recent studies have shown that using weight averaging during pretraining can improve PTQ performance compared with learning-rate decay, suggesting that it might provide …

المؤلفون المشاركون