الباحثون

Yanqi Hao

المنشورات 2

نسخة أولية وصول مفتوح

Offline Guidance, Online Reasoning: Reusing LLM Feedback for Small Language Models

Bohan Zhang, Linan Yue, Weibo Gao وآخرون · 2026

Large language models (LLMs) offer strong reasoning capabilities but are often costly to access through commercial APIs, while small language models (SLMs) are easier to deploy locally yet remain weaker in reasoning. This capability-deployment gap has motivated LLM-SLM collaboration, which aims to improve SLM reasoning …

نسخة أولية وصول مفتوح

G$^2$PTQ: Improving LLM Post-Training Quantization with Generalized Gradient Compensation

Ruikang Liu, Haoli Bai, Yuxuan Sun وآخرون · 2026

Post-training quantization (PTQ) is a practical approach to reducing the memory and computational footprint of large language models (LLMs) without retraining. GPTQ-based methods have become the de facto standard, yet they suffer from two complementary limitations. Methods with local, layer-wise objectives lack global …

المؤلفون المشاركون