الباحثون

Chen Qian

المنشورات 4

نسخة أولية وصول مفتوح

MASBench: Benchmarking LLM-based Multi-Agent Collaboration under Partial Observability

Qizhi Chu, Zekai Yu, Sijie Wen وآخرون · 2026

Large language models (LLMs) have progressively evolved into the core of autonomous agents. Building on this progress, LLM-based multi-agent systems (MAS) coordinate multiple agents into a synergistic team to accomplish complex tasks that exceed the capabilities of individual agents. The effectiveness of such systems d …

نسخة أولية وصول مفتوح

Text-Centric Post-Training for Omni-Modal Reasoning

Ziyang Cheng, Yuhao Wang, Hongcheng Liu وآخرون · 2026

Improving joint audio-visual reasoning in Omni Large Language Models typically incurs substantial data construction and training costs. Our diagnostics reveal multi-hop reasoning difficulties despite correct answers to all corresponding single-hop questions and suggest partial decoupling in the local optimization of pe …

نسخة أولية وصول مفتوح

Rethinking Visual Token Compression for Video Large Language Models: A Simple Yet Strong Baseline

Xiao Zhang, Wang Zeng, Sheng Jin وآخرون · 2026

Video Large Language Models (Video LLMs) have achieved remarkable progress in video understanding, but their inference efficiency is constrained by the large number of visual tokens produced by long videos. Recent video token compression methods increasingly introduce sophisticated strategies for token selection, pruni …

نسخة أولية وصول مفتوح

VIDEAS: Distilling Explicit Action Semantics from Demonstration Videos for World Models via Prior-Guided Simulation

Jianan Wang, Haoquan Zhai, Siyang Zhang وآخرون · 2026

World models learn internal representations of environment dynamics to predict future states, enabling agents to optimize action plans without physical interactions. However, developing world models that genuinely internalize underlying causal physical laws to explicitly reason about action preconditions and subsequent …

المؤلفون المشاركون