Authors

Wei Zhao

Publications 4

Preprint Open access

TReVS: Integrating Textual Relevance and Visual Saliency for Efficient Vision-Language Model Token Pruning

Jing Wang, Zhiping Wu, Dongdong Ren et al. · 2026

Vision-Language Models (VLMs) excel at visual understanding and reasoning but often incur substantial inference costs due to the large number of visual tokens. Recent visual token pruning methods increasingly follow a two-stage paradigm: they first remove visually redundant tokens after the vision encoder and then disc …

Co-authors