Authors

Yinan Wu

Publications 1

Preprint Open access

Quantile Head for Vision-Language-Action Models

Xuan Wang, Yinan Wu, Haoran Duan et al. · 2026

Vision-Language-Action (VLA) models integrate pretrained Vision-Language Models (VLMs) with action heads for robot control. Common action heads have distinct limitations: point regression provides only a point estimate of the action distribution, while standard flow-matching samplers require costly iterative sampling. …

Co-authors