الباحثون

Yu Yao

المنشورات 3

نسخة أولية وصول مفتوح

NeMo-DCR: Bit-Exact Delta-Compressed Refit for Scalable Agentic RL at Trillion-Parameter Scale

Songlin Jiang, Zhiyu Li, Terry Kong وآخرون · 2026

Agentic reinforcement learning (RL) disaggregates training from rollout, so each policy update must reach the rollout clusters before the next batch. Transferring a full 1T checkpoint for such weight synchronization (refit) takes 87.5 min between two AWS regions. Measurements of BF16 training show that about 1% of weig …

نسخة أولية وصول مفتوح

SquidAgent: Parallelize Wisely, Coordinate Efficiently

Yexiong Lin, Shanshan Ye, Yu Yao وآخرون · 2026

LLM-based agents solve complex multi-step tasks, but sequential execution incurs substantial latency. In principle, parallelizing work across multiple agents should yield near-linear speedups. Yet existing parallel multi-agent systems often run slower than a single-agent baseline. We attribute this gap to two hidden co …

نسخة أولية وصول مفتوح

Scalable Minimal-Change Learning for Controllable Image Editing

Shuo Chen, Fengming Huang, Yu Yao وآخرون · 2026

Image editing should change only the attributes specified by an instruction while preserving everything else, yet current methods often make unintended changes. We treat this minimal-change principle as an optimization objective for instruction-based editing. Latent L1 regularization is a poor proxy for output locality …

المؤلفون المشاركون