الباحثون

Wenrui Bao

المنشورات 2

نسخة أولية وصول مفتوح

Recursive Video In-Context Learning for Agentic Robot

Wenrui Bao, Xinxin Liu, Bingxin Xu وآخرون · 2026

LLM agents that orchestrate frozen vision-language-action (VLA) policies improve across episodes through text memory, which records what the agent did but not how the task is done. A demonstration video shows it, but fits poorly into an agent's context. The full video slows every turn, fixed keyframes lose the contact …

نسخة أولية وصول مفتوح

DeltaWAM: Change-Centric Visual Foresight via Delta Tokens for an Efficient World-Action Model

Tianyun Jiang, Wenrui Bao, Bingxin Xu وآخرون · 2026

World-Action Models (WAMs) offer visual foresight for robotic manipulation, but pixel-space models repeatedly reconstruct entire future scenes, incurring high computational cost and spatio-temporal redundancy. In physical manipulation, consecutive frames often share most of their visual context; the changes between the …

المؤلفون المشاركون