الباحثون

Qi Zhu

المنشورات 6

نسخة أولية وصول مفتوح

IntactWorld: Joint World Modeling with Intact Features

Boming Tan, Xiangdong Zhang, Yan Xia وآخرون · 2026

While recent video generation models synthesize highly realistic visuals, they lack a genuine understanding of intrinsic real-world logic. Existing methods attempt to understand the world by internalizing diverse world knowledge, yet constrained by computational overhead or dimensionality alignment, their learning proc …

نسخة أولية وصول مفتوح

UXBench Pro: Benchmarking Personalized User Experience in Multi-Turn Dialogue Interactions

Mengze Hong, Zeyang Lei, Wenbo Shang وآخرون · 2026

Evaluating user experience (UX) with automated computational methods has gained increasing attention, supported by empirical evidence from UXBench. However, binary preference prediction provides limited insight, while relying on a single user-agnostic reward model overlooks the inherent heterogeneity of users, whose ex …

نسخة أولية وصول مفتوح

Beyond Visual Enhancement: Adaptive Multi-Context Steering to Mitigate LVLM Hallucinations

Shuran Ma, JiaLe Li, Yuxin Dong وآخرون · 2026

Hallucination remains a significant challenge in Large Vision-Language Models (LVLMs). Existing training-free methods generally mitigate hallucinations through contrastive decoding or visual enhancement, often increasing the relative influence of visual evidence during generation. This raises a fundamental question: Ca …

نسخة أولية وصول مفتوح

Efficient Multimodal Inference through Adaptive Acquisition and Sequential Fusion

Payal Mohapatra, Haodong Yang, Yueyuan Sui وآخرون · 2026

Multimodal systems often encode every available input, even when a subset suffices for prediction. Adaptive acquisition can reduce this cost by using predictions from incrementally fused evidence to decide which modality to encode next and when to stop. However, sequential fusion makes these predictions order-dependent …

نسخة أولية وصول مفتوح

Source-Learned Reliance for Selective Test-Time Adaptation of Multimodal Time Series

Payal Mohapatra, Yueyuan Sui, Haodong Yang وآخرون · 2026

Multimodal wearable systems must remain reliable when sensor streams become noisy or unavailable. Existing multimodal test-time adaptation (TTA) methods often assess reliability online, but cross-modal agreement can be misleading when sensors measure different physical processes, and evaluating alternative modality con …

نسخة أولية وصول مفتوح

BrainNet Studio: A Unified Toolkit for Brain Network Construction, Intelligent Analysis, and Visualization

Xiwei Zeng, Shengrong Li, Yiheng Liu وآخرون · 2026

Brain networks characterize structural and functional relationships among brain regions and support research on cognition, brain disorders, and brain-computer interfaces. Their time-varying topology and higher-order spatiotemporal dependencies are not adequately represented by conventional static networks. Existing too …

المؤلفون المشاركون