الباحثون

Yuxin Yang

المنشورات 2

نسخة أولية وصول مفتوح

EyeVQA: Benchmarking Ophthalmic Vision-Language Models from Recognition to Spatial Grounding

Gujie Shao, Zixun Xie, Xuechun Xing وآخرون · 2026

Vision-language models (VLMs) have shown increasing potential for medical image understanding, yet their capabilities in ophthalmic imaging remain insufficiently characterized. Existing ophthalmic datasets are typically designed for individual diseases or specialized tasks, making it difficult to systematically evaluat …

نسخة أولية وصول مفتوح

Beyond Reconstruction Error: Analytical and Data-Driven Action Tokenization for Autoregressive Vision-Language-Action Models

Yuxin Yang, Gaohan He, Changxue Guan وآخرون · 2026

Discrete action tokenization is central to autoregressive vision-language-action (VLA) models, yet action representations are often evaluated primarily through reconstruction fidelity. We ask which representation properties actually matter for closed-loop control by comparing fixed analytical, data-driven linear, and n …

المؤلفون المشاركون