نسخة أولية وصول مفتوح
TED:Text-Axis Evidence Decomposition for Prompted Anomaly Localization
CLIP is a powerful vision-language model, but it was not designed for fine-grained defect localization; CLIP-based anomaly detectors therefore adapt it with prompts or lightweight modules to increase defect sensitivity. We show that stronger sensitivity does not necessarily make local evidence reliable: under domain sh …