الباحثون

Yinan Wu

المنشورات 1

نسخة أولية وصول مفتوح

Quantile Head for Vision-Language-Action Models

Xuan Wang, Yinan Wu, Haoran Duan وآخرون · 2026

Vision-Language-Action (VLA) models integrate pretrained Vision-Language Models (VLMs) with action heads for robot control. Common action heads have distinct limitations: point regression provides only a point estimate of the action distribution, while standard flow-matching samplers require costly iterative sampling. …

المؤلفون المشاركون