الباحثون

Mingcong Li

المنشورات 1

نسخة أولية وصول مفتوح

Gaze Prompts: Temporally Dense Human Attention for Vision-Language-Action Fine-Tuning

Yihan Zhou, Rui Yan, Mingcong Li وآخرون · 2026

Vision-Language-Action (VLA) fine-tuning pairs images with actions at every step, yet typically provides only a task-level language instruction, leaving moment-to-moment visual relevance implicit. We introduce \emph{eye-tracker-supervised gaze prompting}, which uses gaze recorded during VR teleoperation to provide fram …

المؤلفون المشاركون