الباحثون

Haobo Yang

المنشورات 1

نسخة أولية وصول مفتوح

VepAgent: Bridging Causal-Transition via Tool-Augmented Reinforcement Learning for Video Event Prediction

Qiutong Chen, Yuchan Guo, Zhenlong Yuan وآخرون · 2026

Multimodal Large Language Models (MLLMs) have demonstrated remarkable potential in video understanding, yet their reliance on retrospective summarization and text-centric priors often limits their ability to bridge unobserved causal transitions when applied to Video Event Prediction (VEP). To address this, we propose V …

المؤلفون المشاركون