الباحثون

Runwei Guan

المنشورات 2

نسخة أولية وصول مفتوح

Answer with Evidence: Consistency-Aware Grounded Visual Question Answering for Roadside Traffic Scenes

Runwei Guan, Rongsheng Hu, Shangshu Chen وآخرون · 2026

Roadside traffic reasoning requires every free-form textual claim to be backed by visual evidence. Existing grounded multimodal large language models (MLLMs) frequently exhibit say-point mismatch, in which the textual answer contradicts the bounding boxes the model localizes. Evaluation metrics that score answers and b …

نسخة أولية وصول مفتوح

RCVLA: 4D Radar-Grounded Semantic Reasoning and Trajectory Arbitration for Autonomous Driving

Lianqing Zheng, Xiaokai Bai, Yixuan Luo وآخرون · 2026

4D radar provides geometric and motion cues that complement visual semantics, but integrating it into vision-language-action (VLA) models requires both radar--language alignment for semantic reasoning and explicit use of radar measurements for trajectory refinement and selection. To support these capabilities, we const …

المؤلفون المشاركون