الباحثون

Lin Zhao

المنشورات 9

نسخة أولية وصول مفتوح

CIRRA: Dual-Level Continual Instruction Reconciliation with Ongoing Execution for Embodied Robot Agents in Interactive Household Tasks

Ci Zhang, Enfu Nan, Arman Akbari وآخرون · 2026

Household robots must accommodate new user instructions while executing ongoing tasks. Existing agents often regenerate or extensively revise the remaining task sequence, introducing plan ambiguity, logical inconsistency, and redundant execution. We formulate continual instruction reconciliation and propose CIRRA (Cont …

نسخة أولية وصول مفتوح

ReCo: Response-Consistent Locomotion with Policy-Aware MPC for Legged Manipulation

Kuankuan Sima, Yichao Gao, Chenxi Gu وآخرون · 2026

Continuous legged manipulation requires accurate end-effector tracking while the base keeps walking. Combining reinforcement learning (RL) with model predictive control (MPC) suits this task: the learned policy provides robust locomotion, while MPC coordinates the base and arm to compensate for tracking errors. However …

نسخة أولية وصول مفتوح

AeroManip-VLA: Scalable Vision-Language-Action Learning for Aerial Manipulation with RL-Generated Demonstrations

Rui Huang, Yanlin Mu, Lidong Li وآخرون · 2026

Aerial manipulators extend robotic manipulation into 3D workspaces that are difficult for ground-based robots to access, creating new opportunities for general-purpose manipulation. However, extending Vision-Language-Action (VLA) models to aerial robots introduces distinct challenges due to the tight coupling between m …

نسخة أولية وصول مفتوح

NeuronEye: Query-Guided Visual Concept Activation for Vision-Language Reasoning

Ruiyu Yan, Bowen Chen, Shaowen Wan وآخرون · 2026

Current vision-language models (VLMs) encode visual information in dense hidden states where object identity, spatial layout, and local attributes are implicitly entangled rather than explicitly disentangled, limiting their ability to isolate and modulate the specific visual evidence required by a given language query. …

نسخة أولية وصول مفتوح

FORGE: Form-Optimal Routing of Grounded Evidence for Frozen LLM Agents

Xi Xiao, Yunbei Zhang, Chen Liu وآخرون · 2026

In agentic AI systems, frozen foundation models are increasingly deployed as closed-weight API endpoints, making downstream adaptation possible only through the inputs and inference procedures surrounding the model. As a result, for each input query, two coupled decisions largely determine both answer quality and token …

نسخة أولية وصول مفتوح

DualGuard: Dual-Mode Quality Control for Logic-Preserving Data Augmentation

Large language models provide a practical way to generate augmented data for logical reasoning at scale, but a larger generation volume does not guarantee semantic, label, or logical reliability. Existing work has improved generation quality through generation constraints, candidate validation, filtering, and feedback- …

نسخة أولية وصول مفتوح

Learning Vision-Based Agile Gap Traversal: Differentiable Simulation with a Warm-Started Critic

Traversing narrow gaps is challenging for autonomous quadrotors, especially when control commands come directly from high-dimensional visual observations. Existing end-to-end methods often rely on behavior cloning or full-rollout backpropagation through time (BPTT) via differentiable simulation, which can limit policy …

المؤلفون المشاركون