الباحثون

Hong Zhang

المنشورات 4

نسخة أولية وصول مفتوح

eRLT: Efficient VLA Reinforcement Learning via Action-Relevant Token Routing

Dehao Huang, Jianbang Liu, Jianpan Gao وآخرون · 2026

Vision-Language-Action (VLA) models provide strong behavioral priors for robotic manipulation, yet efficiently adapting them to downstream tasks remains challenging. Recent work addresses this challenge by adapting frozen VLAs through online reinforcement learning (RL), whose sample efficiency depends on the quality of …

نسخة أولية وصول مفتوح

Function beyond Form: Functional Correspondence for Cross-Embodiment Dexterous Grasp Generation

Bolin Zou, Wenlong Dong, Mu Ai وآخرون · 2026

Cross-embodiment dexterous grasp generation remains challenging because robotic hands differ substantially in geometry, topology, and kinematics. Existing approaches often lack explicit correspondences between structurally different hand regions that play similar functional roles in a grasp, a concept we refer to as fu …

نسخة أولية وصول مفتوح

Text-Vision Synergistic Token Caching: A Training-Free Framework for Efficient Vision-Language-Action Inference

Qianer Li, Chengjie Zhang, Jingwen Chen وآخرون · 2026

Vision-Language-Action (VLA) models enable generalizable robotic control but remain computationally expensive. Token caching provides a training-free, plug-and-play acceleration alternative. However, existing VLA caching does not fully exploit a key inductive bias of VLA models: text-vision synergy, wherein textual sem …

المؤلفون المشاركون