الباحثون

Dave Zhenyu Chen

المنشورات 1

نسخة أولية وصول مفتوح

SpatialSpeak: QA-Native Reconstruction with Local and Global Context for Spatial Chain-of-Thought Reasoning

Yang Cao, Jiaxin Zhang, Dave Zhenyu Chen وآخرون · 2026

Vision-language models (VLMs) can benefit from geometric priors for multi-view spatial reasoning, yet answer-only training does not directly supervise the intermediate geometric estimates and their use in deriving quantitative spatial answers. We hypothesize that spatial chain-of-thought (CoT) supervision becomes more …

المؤلفون المشاركون