الباحثون

Xiyuan Shen

المنشورات 1

نسخة أولية وصول مفتوح

Can Vision-Language Models Analyze Human-Centered Video? Mapping Model Capabilities and Human-AI Collaborative Workflows

Xiyuan Shen, Jiuyang Lyu, Seokhyun Hwang وآخرون · 2026

Video provides a rich record of human behavior, interaction, and situated contexts, offering important evidence for understanding people and conducting human-centered research. As vision-language models (VLMs) become increasingly capable of analyzing video, they offer opportunities to automate this traditionally human- …

المؤلفون المشاركون