الباحثون

Haoli Bai

المنشورات 2

نسخة أولية وصول مفتوح

HARPO: Hallucination-Aware Reinforcement Learning for Faithful and Creative Language Generation

Tiezheng Yu, Yuxin Jiang, Jinpeng Li وآخرون · 2026

Large Language Models (LLMs) are prone to generating hallucinated content, which compromises their reliability in knowledge-intensive tasks. To address this challenge without sacrificing creativity, we propose HARPO, a reinforcement learning framework designed to jointly optimize faithfulness and creativity. HARPO inco …

نسخة أولية وصول مفتوح

G$^2$PTQ: Improving LLM Post-Training Quantization with Generalized Gradient Compensation

Ruikang Liu, Haoli Bai, Yuxuan Sun وآخرون · 2026

Post-training quantization (PTQ) is a practical approach to reducing the memory and computational footprint of large language models (LLMs) without retraining. GPTQ-based methods have become the de facto standard, yet they suffer from two complementary limitations. Methods with local, layer-wise objectives lack global …

المؤلفون المشاركون