الباحثون

Haoran Wen

المنشورات 2

نسخة أولية وصول مفتوح

Streaming-WAM: Action-Conditioned World-Action Model for Asynchronous Robot Manipulation

Xuyao Huang, Yixuan Wang, Zengyao Ye وآخرون · 2026

World action models (WAMs) that use future visual prediction at inference time incur substantial generation costs. Asynchronous execution reduces waiting by overlapping inference with robot motion, but visual predictions used for subsequent action generation must anticipate the effects of actions already scheduled for …

نسخة أولية وصول مفتوح

MachEmbodied-U0: Unified Understanding and Generation Model for Embodied Intelligence

Haoran Wen, Wenfu Wang, Kunsong Shi وآخرون · 2026

General-purpose robot control requires models to understand task intent, identify where to interact, capture how the scene evolves, and generate precise actions. Vision-language-action models provide strong semantic priors but typically do not explicitly model scene dynamics, while world-action models couple visual pre …

المؤلفون المشاركون