الباحثون

Anil Kag

المنشورات 1

نسخة أولية وصول مفتوح

Token-Level Video Reinforcement Learning

Yifan Wang, Gordon Guocheng Qian, Yanyu Li وآخرون · 2026

Reinforcement learning (RL) for video generation usually assigns one scalar reward to an entire sampled video. Yet a video is not uniformly flawed: some visual tokens may already satisfy the prompt, whereas others require correction. A scalar reward cannot localize errors, causing optimization to perturb satisfactory t …

المؤلفون المشاركون