الباحثون

Chloe K. Nobuhara

المنشورات 1

نسخة أولية وصول مفتوح

HeiCo-FOCUS: A Clinically Grounded Dataset for Long-Context Video Understanding

Leon Mayer, Lucas Luttner, Patrick Godau وآخرون · 2026

Recent advances in Vision-Language Models (VLMs) have led to rapid progress in video understanding across a wide range of benchmark tasks. However, existing evaluations largely focus on short-term reasoning, failing to assess a critical capability: maintaining cumulative temporal consistency over extended time horizons …

المؤلفون المشاركون