الباحثون

Yichong Huang

المنشورات 2

نسخة أولية وصول مفتوح

Adaptive Mutual Distillation for Balanced Multi-Task Post-Training of Large Language Models

Baohang Li, Xiaocheng Feng, Yichong Huang وآخرون · 2026

Multi-task post-training of large language models (LLMs) aims to improve performance across tasks with unequal amounts of training data. Existing methods focus primarily on balancing task contributions during single-model training. Different task-balancing strategies can produce models with complementary strengths, cre …

نسخة أولية وصول مفتوح

Collective Bias Mitigation via Model Routing and Collaboration

Mingzhe Du, Luu Anh Tuan, Xiaobao Wu وآخرون · 2026

Large language models (LLMs) are increasingly deployed in public health, finance, and governance, requiring both accuracy and societal value alignment. Despite recent advances, LLMs often perpetuate or amplify bias embedded in their training data, posing challenges to fairness. While self-debiasing encourages an LLM to …

المؤلفون المشاركون