Preprint Open access
PCLM: Small-target localization with frozen CLIP via prototype contrast and local magnification
Small targets occupy few patches in a vision-language encoder, so spatial features often mix object appearance with surrounding content. We propose Prototype Contrast and Local Magnification (PCLM), a support-conditioned localization method that uses a frozen CLIP encoder. Five masked support images per class define fo …