الباحثون

Zekun Zhou

المنشورات 2

نسخة أولية وصول مفتوح

Adaptive Mutual Distillation for Balanced Multi-Task Post-Training of Large Language Models

Baohang Li, Xiaocheng Feng, Yichong Huang وآخرون · 2026

Multi-task post-training of large language models (LLMs) aims to improve performance across tasks with unequal amounts of training data. Existing methods focus primarily on balancing task contributions during single-model training. Different task-balancing strategies can produce models with complementary strengths, cre …

نسخة أولية وصول مفتوح

CRISP: Cultural Reward Modeling for Implicit Situated Propriety

Zekun Yuan, Yangfan Ye, Baohang Li وآخرون · 2026

As large language models (LLMs) are increasingly deployed across countries and regions, the ability to recognize and respond appropriately to diverse cultural contexts becomes increasingly important. However, existing research has largely focused on cultural knowledge or tasks with predefined response spaces, while ope …

المؤلفون المشاركون