الباحثون

Yixiao Huang

المنشورات 2

نسخة أولية وصول مفتوح

How RL Reshapes LLM Reasoning: Transferability, Coverage, and Scaling Laws

Ziheng Cheng, Yixiao Huang, Hanlin Zhu وآخرون · 2026

Recent studies on reinforcement learning (RL) report seemingly conflicting evidence about large language model (LLM) reasoning. Training on mathematics can improve performance in other domains, yet gains in Pass@1 can coincide with lower Pass@$N$ than the base model. This raises a fundamental question: does RL expand a …

نسخة أولية وصول مفتوح

A GHOST in Long-Horizon Agents: Governance Hazard from Overlooked Safety Constraints across Turns

XinPeng Shen, Lan Zhang, Yixiao Huang وآخرون · 2026

Long-horizon agents are now playing an increasingly significant role in assisting humans with complex problem-solving. However, it is exactly their extended interaction history that introduces an underexplored execution-safety concern. Under benign interaction conditions, an agent may execute an action that violates a …

المؤلفون المشاركون