الباحثون

Shengchao Hu

المنشورات 2

نسخة أولية وصول مفتوح

Alignment-Guided Flow Transformer for Efficient Vision-Language-Action Policy Learning

Shengchao Hu, Peng Wang, Qiyang Zhou وآخرون · 2026

Recent advances in Vision-Language-Action (VLA) models point toward general-purpose robotic intelligence by unifying perception, instruction, and control. Despite impressive progress, existing VLA models often adapt poorly due to \emph{tri-modal misalignment} among vision, language, and action, which weakens action gro …

المؤلفون المشاركون