الباحثون

Zhonghua Wang

المنشورات 1

نسخة أولية وصول مفتوح

LoopVL: Recurrent Visual Intelligence

Zhe Qian, Ziyang Gong, Zhongxing Xu وآخرون · 2026

We introduce LoopVL to study whether Loop Transformers can be effectively extended to vision- language models. LoopVL combines Module-Loop and Model-Loop computation to iteratively update a unified vision-language state through shared modules. We train LoopVL from scratch through language pre-training, multimodal train …

المؤلفون المشاركون