الباحثون

Li Shen

المنشورات 8

نسخة أولية وصول مفتوح

Intervention anchors and scientific verification in synthetic vascular predictive representations

Lingsen You, Yujun Guo, Xinyu Zhong وآخرون · 2026

Complete orthogonal predictive coordinates do not by themselves bind a latent direction to a named intervention. We present a mathematical and synthetic audit motivated by vascular device-vessel suitcordance. Capacity-matched least-squares predictors were exactly equivalent under complete fixed output transforms, where …

نسخة أولية وصول مفتوح

Aligning Multimodal Patient Evidence with Biomedical Knowledge Graphs for Clinical LLMs

Jiawen Du, Arshan Ali Khan, Chenhao Zhang وآخرون · 2026

Clinical questions often depend on linking a patient's multimodal evidence to external biomedical knowledge, yet existing predictive systems rarely represent such links explicitly, so they can neither be traced to their evidence sources nor removed to measure their contributions. We present MM-KG (Multimodal Knowledge …

نسخة أولية وصول مفتوح

RoMod: Temporal Routing Modulation via Mixture-of-Experts for Video Anomaly Detection

Chao Huang, Pengfei Wei, Benfeng Wang وآخرون · 2026

Intermediate-layer features from multimodal large language models have shown strong potential for video anomaly detection (VAD), yet the origin of their discriminative power remains unclear. We study this question using sparse mixture-of-experts (MoE) models, whose explicit expert structure and sparse activation make t …

نسخة أولية وصول مفتوح

Pareto-Improving Adversarial Attacks with Primal-Dual Regularization

Yang Dai, Longfei Zhang, Wei Tao وآخرون · 2026

Transferable adversarial attacks are arguably the most practical black-box threat model. Under the same perturbation budget, stronger transfer attacks attain higher attack success rate (ASR), yet their imperceptibility also tends to degrade. Under such a fixed-budget protocol, transferability and imperceptibility there …

نسخة أولية وصول مفتوح

Alignment-Guided Flow Transformer for Efficient Vision-Language-Action Policy Learning

Shengchao Hu, Peng Wang, Qiyang Zhou وآخرون · 2026

Recent advances in Vision-Language-Action (VLA) models point toward general-purpose robotic intelligence by unifying perception, instruction, and control. Despite impressive progress, existing VLA models often adapt poorly due to \emph{tri-modal misalignment} among vision, language, and action, which weakens action gro …

نسخة أولية وصول مفتوح

World Models for Embodied Intelligence: From Plausible to Controllable to Actionable

Nanjie Yao, Hao Wang, Chong Cheng وآخرون · 2026

World models connect perception and decision-making in embodied intelligence by maintaining hidden state, anticipating consequences, comparing interventions, and adapting when execution departs from expectations. Although progress is often measured by visual fidelity, their value lies in improving behavior. Before reac …

المؤلفون المشاركون