الباحثون

Chenhang Cui

المنشورات 4

نسخة أولية وصول مفتوح

BARQ: Balanced Codebook Refinement for Low-Bit LLM Quantization

Chenhang Cui, Xu Xie, Linrui Xu وآخرون · 2026

As large language models (LLMs) grow in parameter count, model storage and parameter memory traffic have become major bottlenecks to efficient deployment. Codebook-based weight quantization reduces these costs, but imbalanced nearest-codeword assignments during fitting can leave some codewords insufficiently updated, l …

نسخة أولية وصول مفتوح

ACTR: Aligning Thoughts and Responses for Multilingual Safety in Reasoning LLMs

Xianhui Zhang, Jian Yu, Chengyu Xie وآخرون · 2026

Ensuring the safety of reasoning large language models (LLMs) across languages is essential for their reliable deployment. However, when exposed to jailbreak attacks in non-high-resource languages, these models may generate unsafe responses even when their reasoning traces identify safety risks. To address this issue, …

المؤلفون المشاركون