الباحثون

Sergey Zakharov

المنشورات 4

نسخة أولية وصول مفتوح

What 30,000 Hours of Ego-centric Video Does Not Teach

Jiahua Dong, Anurag Bagchi, Yash Jangir وآخرون · 2026

World models offer a promising alternative to physics-based simulators, yet remain far from practical deployment. We ask how far scaling ego-centric human video takes them, using a dataset of 30,000 hours spanning over 1,000 scene types and 14,000 contributors. Rather than relying on opaque downstream metrics, we direc …

نسخة أولية وصول مفتوح

MobileVISTA: Generative Data Augmentation for Pose Generalization in Mobile Manipulation

Mobile manipulators such as humanoid robots are increasingly deployed in dynamic, unstructured environments to perform dexterous manipulation tasks. However, end-to-end manipulation policies trained to imitate demonstration data collected from a single robot pose are brittle: even centimeter-scale deviations in robot p …

نسخة أولية وصول مفتوح

BIND: Binding 3D Robot Actions to 2D Image Features

Cameron Smith, Arsh Tangri, Vitor Guizilini وآخرون · 2026

We introduce BIND, a new action representation for visuomotor robot policies that binds 3D robot actions to their corresponding 2D image features, yielding strong data efficiency gains and robustness to out-of-distribution object positions and camera viewpoints. The action heads of current robot policies are typically …

نسخة أولية وصول مفتوح

H2RBench: A Real-to-Sim Benchmark for Evaluating Human-to-Robot Transfer

Chuyang Xiao, Haotian Zhan, Sriram Krishna وآخرون · 2026

Learning robot manipulation policies from human video demonstrations constitutes a promising avenue for scalable robot learning. However, comparing different human-to-robot (H2R) transfer methods remains challenging, as existing approaches are evaluated under different settings, including differing task suites, scene l …

المؤلفون المشاركون