الباحثون

Dahua Lin

المنشورات 3

نسخة أولية وصول مفتوح

RollVerify: Bridging Efficiency and Accuracy in Long-Tail Rollout Reinforcement Learning

Yongqiang Yao, Jinru Tan, Kaihuan Liang وآخرون · 2026

Reinforcement learning is crucial for improving large language models' reasoning and generalization. It relies on massive rollouts whose lengths become increasingly long-tailed as context windows grow. In on-policy training, these long-tail rollouts can result in GPU bubbles, reducing system utilization and limiting RL …

نسخة أولية وصول مفتوح

Looped Diffusion Transformer

Yong Xien Chng, Tianyi Chen, Wenwen Tong وآخرون · 2026

Improving text-to-image models has traditionally relied on increasing model size or the number of denoising steps. In this work, we explore an alternative way to scale computation by repeatedly running shared Transformer blocks within each denoising step, effectively increasing computational depth while keeping the par …

نسخة أولية وصول مفتوح

Skel-WAM: A Hand-Skeleton-Conditioned World Action Model for Human-to-Robot Manipulation Transfer

Zetao Cai, Yaping Li, Yiqun Wang وآخرون · 2026

Robot demonstrations are expensive to collect and often provide limited distributional coverage of task variations. Human videos offer a low-cost source of complementary manipulation experience, but learning from them requires bridging embodiment gaps in visual appearance and action spaces. We introduce Skel-WAM, a wor …

المؤلفون المشاركون