الباحثون

Zixuan Wang

المنشورات 6

نسخة أولية وصول مفتوح

RESETTLE: Robotic Recovery through Disagreement-Triggered Retrieval and Efficient Corrective Control

Yuxin Chen, Senqiao Yang, Zixuan Wang وآخرون · 2026

Reliable robotic manipulation requires timely intervention to correct emerging deviations and restore progress after execution errors. However, recovery methods based on repeated vision-language reasoning or iterative online optimization can incur substantial latency, delaying intervention. To address these challenges, …

نسخة أولية وصول مفتوح

Distilling Routed 3D Privilege for Spatial Reasoning in Vision-Language Models

Hongxing Li, Yixin Li, Dingming Li وآخرون · 2026

Spatial reasoning remains a persistent weakness of vision-language models (VLMs), because RGB inputs do not directly provide geometric evidence. Existing remedies either inject 3D into the model at inference, paying architecture and latency costs, or train with outcome rewards that supervise only the final answer. Spat …

نسخة أولية وصول مفتوح

Beyond Corrected Memory: Execution Consistency in Multi-Agent Systems

Zhe Yu, Zixuan Wang, Peidong Wang وآخرون · 2026

Shared memory coordinates agents' actions, but correct records do not establish that those actions satisfy task requirements. Memory governance and failure diagnosis regulate or inspect recorded information; they do not by themselves establish whether it is sufficient to judge task duties. We define execution consisten …

نسخة أولية وصول مفتوح

Self-Reflection Fine-Tuning: Enhancing Agent Security against Prompt Injection Attacks from Failure Experience

Zixuan Wang, Hao Li, Fengyu Gao وآخرون · 2026

Large language model (LLM) agents are increasingly deployed in tool-augmented environments, but their reliance on external inputs makes them highly vulnerable to prompt injection attacks that can hijack task objectives. Existing safety alignment methods rely on static expert trajectories or preference optimization, lim …

نسخة أولية وصول مفتوح

Harness Learning Enables Generalizable Test-Time Adaptation

Alvin Zhang, Xuecheng Liu, Zixuan Wang وآخرون · 2026

A language-model agent is jointly defined by its model and its harness, the executable program that organizes model calls, tool use, and information flow. Because different tasks call for different ways of organizing these operations, the harness needs to be adapted using feedback from the task at hand. We introduce ha …

نسخة أولية وصول مفتوح

Mind2Dialogue: Training Human-Aware Language Models by Simulating User Mental States

Zixuan Wang, Yufan Zhou, Jinzhou Tang وآخرون · 2026

As language models become more capable, long-term collaboration in learning, reasoning, and decision-making calls for a deeper understanding of the people they serve. Yet training such human-aware language models faces a fundamental supervision gap because current datasets for LLM assistant training contain few if any …

المؤلفون المشاركون