الباحثون

Li Fei-Fei

المنشورات 6

نسخة أولية وصول مفتوح

Cross-Embodiment Robot Foundation World Models with Latent Actions

The diversity of robot embodiments and action spaces makes it challenging to build robot world models that generalize across different embodiments. We introduce the Latent Action-Conditioned Robot World Model (LAC-WM), which operates within a learned unified latent action space shared across diverse embodiments. This u …

نسخة أولية وصول مفتوح

OpenWAM: An Open Framework for Composable World-Action Models

Heng Yu, David D. Yuan, Juze Zhang وآخرون · 2026

World-action models (WAMs) couple future prediction with robot control, yet existing systems often vary the video backbone, interaction structure, supervision, and inference procedure simultaneously, making their design choices difficult to compare. We introduce OPENWAM, an open world-action modeling framework built ar …

نسخة أولية وصول مفتوح

LoGo: Local-Global Rewards for Consistent Long-Horizon Video Generation

Ziqi Ma, Shreya Sharma, Mohamed El Banani وآخرون · 2026

Camera-controlled video models are rapidly advancing toward long generation horizons and complex camera control. A key failure mode is 3D inconsistency: as the camera moves, objects lose permanence and scene structures shift. Existing post-training techniques, which assign a single scalar reward to the entire generatio …

نسخة أولية وصول مفتوح

T$^2$Mem: Learning Test-Time Memory for Robotics

Yize Liu, Huang Huang, Yining Hong وآخرون · 2026

Memory-dependent robotic manipulation requires policies to use information that is no longer available in the current observation. Retaining history alone is insufficient: memory must preserve information that supports future actions. One challenge is whether a memory-free foundation model can learn to retain and use h …

نسخة أولية وصول مفتوح

DexAgent: An Agentic Human2Sim2Robot Framework for Dexterous Manipulation with Self-Evolving Tool Library

Youhui Wang, Yunzhu Li, Li Fei-Fei وآخرون · 2026

Human videos offer a scalable source of demonstrations for dexterous robot manipulation. However, existing human-to-simulation-to-robot (Human2Sim2Robot) pipelines rely on predefined procedures that struggle to accommodate diverse object properties and interactions, particularly those involving articulated and deformab …

المؤلفون المشاركون