الباحثون

Ziyuan Yang

المنشورات 5

نسخة أولية وصول مفتوح

Attributing HOW, Not Just WHICH: Counterfactual Response Trajectories for Diffusion Models

Haoqian Zhang, Ziyuan Yang, Zerui Shao وآخرون · 2026

Diffusion models have achieved remarkable success in image generation, yet tracing their outputs to individual training examples remains challenging. Existing attribution methods often compress factor-specific effects into scalar responses, making distinct internal changes indistinguishable. This is particularly limiti …

نسخة أولية وصول مفتوح

Reading, Not Manipulating: Leveraging Router Logits for Multimodal Safety in MoE Vision-Language Models

Ziyuan Yang, Wenxuan Ding, Shangbin Feng وآخرون · 2026

Vision-language models (VLMs) face compositional safety risks where harmful intent emerges from the interaction between visual and textual inputs. As mixture-of-experts (MoE) VLMs become increasingly common, recent work has explored various safety interventions, including prompting, supervised fine-tuning, and routing- …

نسخة أولية وصول مفتوح

You're Hired: Strategic Model Selection for LLM Collaboration

Zongwan Cao, Ziyuan Yang, Shangbin Feng وآخرون · 2026

While multi-agent and model collaboration algorithms gain traction to combine the strengths of diverse Large Language Models (LLMs), existing systems remain bottlenecked on pre-defined and hand-crafted model pools. In this work, we investigate the problem of model selection in multi-LLM systems. We propose and systemat …

نسخة أولية وصول مفتوح

Save Your Saturated Data: Learning Beyond Reward Saturation in Group-Based RL

Ziyuan Yang, Yike Wang, Shangbin Feng وآخرون · 2026

Group-relative reinforcement learning (RL) relies on reward variation among sampled responses to estimate informative relative advantages. As language models become increasingly capable, existing training data can become reward-saturated: all sampled responses to the same problem might receive equally high rewards, whe …

المؤلفون المشاركون