الباحثون

Ahmadreza Jeddi

المنشورات 1

نسخة أولية وصول مفتوح

SCOPD: Sparse-Context On-Policy Self-Distillation for Efficient Vision-Language Models

Ahmadreza Jeddi, Enming Zhang, Jasper Gerigk وآخرون · 2026

Reasoning vision-language models (VLMs) process images and videos as long sequences of visual tokens, making inference expensive. Training-free token pruning reduces this cost, but aggressive compression can sharply degrade performance, often attributed to irreversible loss of task-relevant visual information. We show …

المؤلفون المشاركون