Authors

Yuhan Sun

Publications 3

Preprint Open access

Visual sensitivity is not claim retractability: persistence-aware credit assignment for multimodal reinforcement learning

Zhongan Bi, Kepeng Lin, Xuanang Gao et al. · 2026

Reinforcement Learning with Verifiable Rewards (RLVR) has been extended to Large Vision-Language Models (LVLMs), and perception-aware methods further encourage policies to rely on visual evidence. Yet relying on the image does not guarantee that visual claims are supported by it. Before RL training, 27.81% of the corre …

Co-authors