الباحثون

Chunyu Zou

المنشورات 2

نسخة أولية وصول مفتوح

D$^2$-VLA: Dual-Memory Dual-Frequency Vision-Language-Action Model For Long Dynamic Manipulation

Zijian Ye, Chengqi Wei, Wei Huang وآخرون · 2026

Long-horizon manipulation requires robots to remember cues that are no longer in view while responding to moving objects. Yet vision-language-action (VLA) policies often rely on the latest observation, and refreshing their visual context typically requires another costly vision-language model (VLM) pass. We present D$^ …

نسخة أولية وصول مفتوح

Proxy2World: Learning to Generate Worlds From Lightweight Proxies without Seeing Them

Hongli Xu, Weilong Yan, Anbang Wang وآخرون · 2026

Lightweight scene proxies let creators control scene layout and motion while leaving room for imagination in appearance, lighting, and visual effects. However, a suitable proxy is not uniquely defined, making paired proxy-video data difficult to construct automatically at scale. We present Proxy2World, a controllable w …

المؤلفون المشاركون