الباحثون

Jun Liu

المنشورات 7

نسخة أولية وصول مفتوح

CVIF: A Criticality-Driven Visual Intervention Framework for Geometric Diagram Understanding in MLLMs

Jiahui Kang, Bifan Wei, Lingling Zhang وآخرون · 2026

Despite significant progress in visual tasks by Multimodal Large Language Models (MLLMs), geometric diagram understanding remains challenging due to the presence of sparse visual cues and ambiguous symbol-primitive associations. MLLMs may therefore rely on textual priors, producing interpretations that conflict with vi …

نسخة أولية وصول مفتوح

TAME:Topology-Aware Text-Driven Motion Editing across Heterogeneous Humanoid Skeletons

Qichen Zheng, Siyuan Yang, Chong Wang وآخرون · 2026

Text-driven motion editing modifies an existing motion sequence according to a text instruction while preserving the content of the source motion. Existing methods are typically built for a single, fixed skeletal topology, which limits their use in animation pipelines where characters differ in joint count and skeletal …

نسخة أولية وصول مفتوح

Does Learning to Predict the World Help Agents Act? Auditing World-Model Post-Training

Xinyu Che, Hang Yan, Yanchen Liu وآخرون · 2026

Predicting how an environment will change before acting is a natural route to better decision making for agents. Recent post-training methods therefore require agents to predict the next observation and turn that prediction into a reward or a direct supervision signal, which is called world model. Existing next-observa …

نسخة أولية وصول مفتوح

Ruby-ASR: Evidence-Preserving Supervision for Joint Orthographic and Lexical-Reading Recognition

Hao Shi, Yun Liu, Xuehao Yang وآخرون · 2026

Conventional Japanese automatic speech recognition (ASR) is supervised by an orthographic transcript, although the same written form can correspond to different lexical readings realized in speech. Such utterances receive an identical target, so their reading distinction is absent from the supervision interface and can …

نسخة أولية وصول مفتوح

Geometric Shortcuts for Complex Trunk Postures: Dual-Helicity Coupling Enables Low-Dimensional Control

Huishi Huang, Danlu Chen, Matteo Lo Preti وآخرون · 2026

How do elephant trunks generate complex postures without relying solely on fine segmental activation? We propose that part of this complexity arises from a low-dimensional geometric shortcut: dual-helicity coupling between opposite-handed oblique muscles. In a simplified soft-robotic prototype, varying only two geometr …

المؤلفون المشاركون