الباحثون

Xiantao Zhang

المنشورات 2

نسخة أولية وصول مفتوح

SPIDER: Multi-Layer Semantic Token Pruning and Adaptive Sub-Layer Skipping in Multimodal Large Language Models

Tianxiang Chen, Zhentao Tan, Zi Ye وآخرون · 2026

Multimodal Large Language Models face significant efficiency challenges that stem from two distinct yet coupled sources: data redundancy and computational redundancy. While most methods focus on data redundancy by pruning visual tokens from the output of the visual encoder or computing redundancy in LLM decoders using …

نسخة أولية وصول مفتوح

Dense to MoE Adaptation for Compact Vision Language Action Policies

Muchun Niu, Shuang Chen, Yuzhou Wu وآخرون · 2026

Vision language action (VLA) policies continue to grow in parameter count, making deployment on resource-constrained robot platforms difficult. The central goal is to reduce the number of LLM-side parameters retained in the deployed policy while preserving downstream task performance. Our approach, AdaDE, adapts select …

المؤلفون المشاركون