الباحثون

Vitor Guizilini

المنشورات 3

نسخة أولية وصول مفتوح

BIND: Binding 3D Robot Actions to 2D Image Features

Cameron Smith, Arsh Tangri, Vitor Guizilini وآخرون · 2026

We introduce BIND, a new action representation for visuomotor robot policies that binds 3D robot actions to their corresponding 2D image features, yielding strong data efficiency gains and robustness to out-of-distribution object positions and camera viewpoints. The action heads of current robot policies are typically …

نسخة أولية وصول مفتوح

Rolling-WAM: World Action Models with Rolling Imagination

Yinghua Zhou, Junjie Ye, Yiqi Zhao وآخرون · 2026

World Action Models (WAMs) couple action generation with future visual prediction for robotic manipulation. However, completing the joint video-action denoising process at each replanning cycle incurs substantial latency, delaying action updates and limiting closed-loop responsiveness. We present Rolling-WAM, a formula …

المؤلفون المشاركون