الباحثون

Chenyuan Wang

المنشورات 2

نسخة أولية وصول مفتوح

UltraG-Bench: A Multi-task Benchmark for assessing Large Vision-Language Models on Pixel-level Evidence Grounding in Ultrasound

Quanhao Zhu, Bo Xu, Rui Lin وآخرون · 2026

Ultrasound is one of the most widely used medical imaging modalities, and recent large vision-language models(VLMs) have shown increasing capabilities in ultrasound image understanding. However, these models fail to provide pixel-level visual evidence aligned with their semantic predictions, and their fine-grained grou …

نسخة أولية وصول مفتوح

Learn Before You Judge: Progressive Knowledge-to-Decision Alignment for Explainable Hateful Meme Detection

Bo Xu, Chenyuan Wang, Xinyu Chen وآخرون · 2026

Hateful memes spread abusive content through implicit interactions between images and text, posing serious threats to the safety of online communities. In recent years, multimodal large language models have been widely used for hateful meme detection and are increasingly adopted to generate explainable detection result …

المؤلفون المشاركون