Authors

Jungong Han

Publications 3

Preprint Open access

LoopVL: Recurrent Visual Intelligence

Zhe Qian, Ziyang Gong, Zhongxing Xu et al. · 2026

We introduce LoopVL to study whether Loop Transformers can be effectively extended to vision- language models. LoopVL combines Module-Loop and Model-Loop computation to iteratively update a unified vision-language state through shared modules. We train LoopVL from scratch through language pre-training, multimodal train …

Preprint Open access

Quantile Head for Vision-Language-Action Models

Xuan Wang, Yinan Wu, Haoran Duan et al. · 2026

Vision-Language-Action (VLA) models integrate pretrained Vision-Language Models (VLMs) with action heads for robot control. Common action heads have distinct limitations: point regression provides only a point estimate of the action distribution, while standard flow-matching samplers require costly iterative sampling. …

Co-authors