الباحثون

Jiancheng Yang

المنشورات 3

نسخة أولية وصول مفتوح

Readout Blindness: VLM Scores Miss the Spatial Direction Their Frozen Encoders Retain

Guangyuan Li, Tianming Du, Yan Jiang وآخرون · 2026

CLIP-like vision-language models remain a cornerstone of multimodal systems, yet their scores stay near chance on directed spatial relations, such as whether one object is left of another. We call this failure readout blindness and analyze, theoretically and empirically, why deployed scores miss the direction: when sco …

نسخة أولية وصول مفتوح

NeurDuo-EEG: A Long-Sequence EEG Foundation Model with Persistent State and Explicit Memory

Yifan Wang, Haiping Liu, Yang Cui وآخرون · 2026

Electroencephalography (EEG) is recorded continuously over hours, with relevant dynamics spanning timescales from milliseconds to hours. Most EEG foundation models nevertheless process fixed windows independently, limiting their ability to capture information encoded in long-timescale dynamics. State-space architecture …

نسخة أولية وصول مفتوح

What Makes a Good Medical Image Tokenizer? Rethinking Reconstruction and Generation in Medical Image Tokenization

Latent diffusion models now dominate medical image generation, and every such pipeline rests on a \emph{tokenizer} that compresses images into the latent codes for image generation to operate on. Thereby, the tokenizer choice bounds every downstream task from reconstruction fidelity and generation quality to the repres …

المؤلفون المشاركون