الباحثون

Jie Wang

المنشورات 6

نسخة أولية وصول مفتوح

OmniRoute: Mapping Temporal Semantic Evidence to Audio-Visual Token Budgets for Efficient Omnimodal Large Language Models

Yuchen Deng, Zidang Cai, Feidiao Yang وآخرون · 2026

Omnimodal large language models (Omni-LLMs) encode audio and visual streams into temporally interleaved token sequences for multimodal reasoning. However, processing long audio-visual token sequences incurs substantial prefill costs. Existing compression methods have made progress, but often overlook temporal changes i …

نسخة أولية وصول مفتوح

CogWAM: Aligning Semantic Cognition with World Action Modeling via Event-Driven Interfaces

Sen Wang, Liu Liu, Xinjiang Wang وآخرون · 2026

Robot policies increasingly incorporate semantic reasoning and future-world prediction, yet combining these capabilities does not guarantee that local predictions and actions remain aligned with task progress. We introduce CogWAM, a cognition-guided world-action model that establishes an explicit semantic interface bet …

نسخة أولية وصول مفتوح

ThinkV2V: Unleashing the Reasoning Capability of MLLMs for Instruction-Guided Video Editing

Donghao Zhou, Haoyang He, Fan Zhang وآخرون · 2026

Instruction-guided video editing has made significant progress, yet existing methods use multimodal large language models (MLLMs) primarily as semantic encoders, so they often fall short in working with implicit edits that require causal or semantic reasoning. To bridge this fundamental gap in video editing, we propose …

نسخة أولية وصول مفتوح

RoboCafé in the Open: Interaction Continuity in Long-Term Public Human-Robot Interaction

As robots remain in public spaces over extended periods, they must maintain interaction continuity by preserving and correctly applying context as people, encounters, and circumstances change. To study interaction continuity in long-term public human-robot interactions, we developed RoboCafé, an autonomous conversation …

نسخة أولية وصول مفتوح

Cosserat Modeling of Trimmed Helicoid Soft Arms with a Separated-Section Constitutive Law

Zhihang Qin, Linxin Hou, Zeyu Zhong وآخرون · 2026

Cosserat rod models for soft robots usually construct sectional stiffness by summing material properties over a common cross-section. This assumption becomes inaccurate for trimmed helicoid arms, where load-bearing helix domains are separated and connected only through sparse fused crossings. This paper formulates a se …

المؤلفون المشاركون