الباحثون

Shuang Chen

المنشورات 2

نسخة أولية وصول مفتوح

Behavior Pack Optimization for Video MLLM Post-Training

Zhaolu Kang, Shiyu Liu, Tailong Luo وآخرون · 2026

Video multimodal large language models (MLLMs) keep climbing video question answering benchmarks, yet shuffling the frames, masking the segment that supports the answer, or occluding the target object barely changes their predictions. The accuracy rests on appearance and language priors, not on the temporal evidence th …

نسخة أولية وصول مفتوح

Dense to MoE Adaptation for Compact Vision Language Action Policies

Muchun Niu, Shuang Chen, Yuzhou Wu وآخرون · 2026

Vision language action (VLA) policies continue to grow in parameter count, making deployment on resource-constrained robot platforms difficult. The central goal is to reduce the number of LLM-side parameters retained in the deployed policy while preserving downstream task performance. Our approach, AdaDE, adapts select …

المؤلفون المشاركون