نسخة أولية وصول مفتوح
Video-HopChain: Multi-Hop Questions and Confidence-Gated Exploration for Video Reasoning Models
HopChain has shown on still images that multi-hop data synthesis improves vision-language reasoning, because long chain-of-thought reasoning exposes errors that compound across steps, while most data used for reinforcement learning with verifiable rewards (RLVR) rarely demands a chain of visual evidence, so these weakn …