الباحثون

Xinyi Zhuang

المنشورات 1

نسخة أولية وصول مفتوح

Representation Dynamics Reveal Semantic Saliency and Similarity for Visual Token Pruning in MLLMs

Weixuan Li, Zikun Zhou, Xinyi Zhuang وآخرون · 2026

Multimodal large language models (MLLMs) incur high inference latency from long visual token sequences. Existing pruning methods commonly use attention maps or output features to estimate token importance or redundancy. Several recent approaches also exploit representation changes, but when and how these changes reflec …

المؤلفون المشاركون