الباحثون

Xiaotong Li

المنشورات 3

نسخة أولية وصول مفتوح

EgoExo-Next:Benchmarking Vision-Language Models on Visual-Option Next-State and Cross-View Reasoning

Yutong Li, Molin Wang, Xiaotong Li وآخرون · 2026

Vision-language models (VLMs) are increasingly evaluated for egocentric and cross-view video reasoning, yet existing benchmarks largely focus on semantic event understanding, temporal relations, or correspondence between already observed views, leaving their ability to reason directly about future visual states underex …

نسخة أولية وصول مفتوح

Track2Art: Articulated Object Model Recovery with Visual-Geometric Track Representations

Xiaotong Li, Yixiong Jing, Junsheng Ding وآخرون · 2026

Understanding articulated objects is fundamental for robotic interaction, requiring accurate rigid-part discovery and the recovery of their kinematic relations. Existing approaches often treat articulation as a by-product of reconstructed geometry or recover it through per-instance optimization. We instead build on the …

المؤلفون المشاركون