الباحثون

Xijie Huang

المنشورات 3

نسخة أولية وصول مفتوح

DiffWAM: A Fast and Efficient Navigation World Action Model

Mo Zhu, Yuze Wu, Xijie Huang وآخرون · 2026

Pretrained video foundation models encode rich semantic and spatiotemporal priors for embodied navigation, yet converting these priors into UAV motion typically requires expensive future-video synthesis and geometric reconstruction. We investigate whether the motion implicit in future visual prediction can instead be r …

نسخة أولية وصول مفتوح

NavGen: Visual Generative Models as a Scalable Data Engine for Embodied 3D Navigation

Xijie Huang, Yongyang Wan, Chengbin Dong وآخرون · 2026

General-purpose robot models increasingly rely on large and diverse datasets. For embodied 3D navigation, however, existing data sources face a fundamental trade-off: simulated data can be generated at scale but often suffer from the visual sim-to-real gap, whereas real-world flight data provide realistic observations …

نسخة أولية وصول مفتوح

TADreamer: Zero-Shot Language-Guided 3D Navigation for Terrestrial-Aerial Bimodal Robots via Video Imagination

Xiangyu Li, Tiancheng Lai, Xijie Huang وآخرون · 2026

Language-guided navigation for terrestrial-aerial bimodal robots requires selecting routes and locomotion modes that match scene context and task intent. Generated videos can represent such motion sequences, but recovering metrically consistent navigation references from them is challenging because of scale ambiguity a …

المؤلفون المشاركون