الباحثون

Guangwen Yang

المنشورات 2

نسخة أولية وصول مفتوح

MASCRDM: Multi-Agent System for Compliance Risk Detection and Mitigation in Training Process of Large Language Models

Yan Zhang, Chuming Wei, Ruien Li وآخرون · 2026

Large Language Models (LLMs) have been applied in various fields. However, ensuring compliance and safety of LLMs, such as avoiding discrimination and bias, still remains a challenge. Current efforts mainly focus on detecting and filtering inputs and outputs of the trained models, rather than studying the intrinsic arc …

نسخة أولية وصول مفتوح

AutoLoCo: Communication Efficient Distributed LLM Training via Adaptive Synchronization

Pengyu He, Yan Zhang, Ruien Li وآخرون · 2026

The pre-training of Large Language Models (LLMs) is increasingly conducted across multiple data centers. As training scales to a larger number of accelerators, the fraction of time spent on computation decreases, while the fraction spent on communication increases. Therefore, frequent synchronization becomes a growing …

المؤلفون المشاركون