الباحثون

Pengfei Wei

المنشورات 3

نسخة أولية وصول مفتوح

Representation--Behavior Alignment for Explainable Weakly-Supervised Video Anomaly Detection

Chao Huang, Pengfei Wei, Kaige Li وآخرون · 2026

Multimodal Large Language Models (MLLMs) provide a natural way to make video anomaly detection more explainable. However, their final decisions do not always fully use the discriminative information contained in their hidden states, an issue we refer to as representation--behavior misalignment. We decompose this gap in …

نسخة أولية وصول مفتوح

RoMod: Temporal Routing Modulation via Mixture-of-Experts for Video Anomaly Detection

Chao Huang, Pengfei Wei, Benfeng Wang وآخرون · 2026

Intermediate-layer features from multimodal large language models have shown strong potential for video anomaly detection (VAD), yet the origin of their discriminative power remains unclear. We study this question using sparse mixture-of-experts (MoE) models, whose explicit expert structure and sparse activation make t …

نسخة أولية وصول مفتوح

VisionPsy-Nano: Improving Accuracy, Efficiency, and Reliability in On-Device Vision-Language Models

Sub-billion-parameter Vision-Language Models are increasingly viable for on-device deployment, yet compact model size alone does not guarantee usability. On a phone, such a model can still require more than two minutes to produce its first token. On-device usability depends on three axes: accuracy, efficiency, and beha …

المؤلفون المشاركون