الباحثون

Hengyi Zhang

المنشورات 2

نسخة أولية وصول مفتوح

SEPAL: Separated Expert Pairs with Answer-Level Fusion for Reliable LLM Collaboration

Weijie Ren, Yanwen Zhang, Hao Li وآخرون · 2026

Multi-agent collaboration lets large language models (LLMs) improve question answering through deliberation and feedback. Yet shared discussion couples correction with exposure to the same mistakes, which can erode the diversity needed for voting. Self-consistency offers sampling diversity without feedback, while singl …

نسخة أولية وصول مفتوح

OPSRD: On-Policy Self-Role Distillation

Weijie Ren, Yanwen Zhang, Hao Li وآخرون · 2026

Role prompting elicits specialized behavior from large language models through an expert identity, offering a lightweight way to guide reasoning on demanding tasks. However, evaluating or distilling complete role-prompted answers can miss useful next-token preferences when the sampled solution remains incorrect. Transf …

المؤلفون المشاركون