الباحثون

Junchi Yan

المنشورات 5

نسخة أولية وصول مفتوح

Reasoning-Informed Visual Editing

Xue Yang, Peiyuan Zhang, Yilun Zhu وآخرون · 2026

Large Multi-modality Models (LMMs) have made significant progress in visual understanding and generation, but still face challenges in visual editing, particularly in following complex instructions, preserving appearance consistency, and supporting flexible input formats. To study this gap, we introduce RISEBench, the …

نسخة أولية وصول مفتوح

ChainLoRA: Geometry-Preserving Task Vector Merging for Continual Learning in LLMs

Hang Yin, Haozhe Wang, Yuhua Luo وآخرون · 2026

Continual parameter-efficient fine-tuning for large language models (LLMs) must balance retention of previously acquired knowledge, adaptation to new tasks, and strict parameter budgets. We present \textbf{ChainLoRA}, a replay-free continual merging framework built on chain-updated task-vector geometry. From a paramete …

نسخة أولية وصول مفتوح

LoopVL: Recurrent Visual Intelligence

Zhe Qian, Ziyang Gong, Zhongxing Xu وآخرون · 2026

We introduce LoopVL to study whether Loop Transformers can be effectively extended to vision- language models. LoopVL combines Module-Loop and Model-Loop computation to iteratively update a unified vision-language state through shared modules. We train LoopVL from scratch through language pre-training, multimodal train …

نسخة أولية وصول مفتوح

Bench2Dex: Benchmarking Visuo-Tactile Bimanual Dexterous Manipulation Across Dexterous Hands

Zhenjie Yang, Yideng Zhang, Dongjie Zhang وآخرون · 2026

Tactile sensing provides contact information that can be difficult to infer from vision alone, but tactile hardware for dexterous hands has not converged to a common design. Dexterous hands differ in finger structure, contact surfaces, and sensor layouts, while simulated tactile signals still differ from measurements p …

المؤلفون المشاركون