الباحثون

Xiaochen Yuan

المنشورات 1

نسخة أولية وصول مفتوح

ATI-VLA: Action-Centric Predictive Vision-Language-Action Models via Actionable Alignment Then Adaptive Injection

Yijie Zhu, Rui Shao, Jie He وآخرون · 2026

Predictive Vision-Language-Action (VLA) models aim to improve robotic manipulation via future observation or world dynamics forecasting. However, existing approaches often fail to realize this potential and underperform direct action prediction models. We argue that these limitations stem from modality misalignment bet …

المؤلفون المشاركون