الباحثون

Yunfan Lou

المنشورات 2

نسخة أولية وصول مفتوح

Token-World: World Modeling in Vision-Language Model Token Space for Robot Manipulation

Chuyao Fu, Xiaowei Chi, Yuhan Rui وآخرون · 2026

A common approach to world-model simulation for vision-language-action (VLA) systems is to predict future RGB observations and then re-encode them into policy inputs, introducing an indirect interface between simulation and downstream policy execution. We instead investigate whether world dynamics can be modeled in a c …

نسخة أولية وصول مفتوح

RoboFL: Federated Expert Assembly for World Action Models

Rongyu Zhang, Ruizhi Fan, Yunfan Lou وآخرون · 2026

Vision-language-action and world-action models are increasingly popular, yet remain bottlenecked by physical interaction data that is scarce, institutionally siloed, and task-heterogeneous. A natural federated solution is to let each client adapt a shared foundation model through parameter-efficient fine-tuning, avoidi …

المؤلفون المشاركون