الباحثون

Zeyu Zhang

المنشورات 12

نسخة أولية وصول مفتوح

RFPO: Rectified Flow Policy Optimization for Embodied Control

Ting Huang, Lisiyu Pan, Haoyu Wang وآخرون · 2026

Flow-based policies provide an expressive framework for continuous robot control, but their iterative ODE integration incurs substantial inference cost. Naively reducing the integration budget can severely degrade control, since policies optimized under full-step execution are not explicitly constrained to remain relia …

نسخة أولية وصول مفتوح

PhysEvo: Astra Can Act, Let It

Wenqing Tian, Zeyu Zhang, Zhaocheng Liu وآخرون · 2026

Astra can act, yet reliable manipulation depends on the system through which it observes and controls the world. We introduce PhysEvo, a framework for physical recursive self-improvement (RSI) around a single frozen model. A task agent executes robot tasks; a meta-agent uses the resulting trajectories to diagnose failu …

نسخة أولية وصول مفتوح

PWM: Personalized World Models with Online Reinforcement Learning

Zhexin Lou, Guancheng Lu, Zeyu Zhang وآخرون · 2026

Pretrained world models can generate diverse environments, yet users often want to explore a particular scene specified by their own video. This requires learning the scene's visual identity while retaining the quality of action-conditioned generation. We introduce Personalized World Models (PWM), a framework for custo …

نسخة أولية وصول مفتوح

Code2Games: Enabling Coding Agents for Gaming World Generation

Wei Wu, Ziyang Xu, Zeyu Zhang وآخرون · 2026

Generating a high-quality gaming world from a natural-language game intent requires joint reasoning about scene structure, spatial layout, gameplay objectives, interactive entities, and executable gameplay logic. Existing coding agents can generate individual assets, scenes, or scripts, but often struggle to maintain c …

نسخة أولية وصول مفتوح

False Frontiers: Diagnosing and Mitigating Co-Cheating in Self-Evolving Search Agents

Meijia Chen, Hao Li, Zheng Lu وآخرون · 2026

Self-evolving search agents build their own training curricula by jointly optimizing a proposer that generates questions and a solver that answers them. This closed loop introduces a failure mode we call co-cheating: the proposer and solver increasingly agree on shared errors, so internal reward improves without a matc …

نسخة أولية وصول مفتوح

Zero2Repo: Can Coding Agents Build Repositories from Scratch?

Pei Yang, Tianyu Shi, Yuhang Yao وآخرون · 2026

Coding agents are increasingly asked to build software rather than patch it, yet benchmarks for from-scratch repository construction are mostly limited to a single language and depend on manually curated tasks. We introduce Zero2Repo, a benchmark in which an agent receives a product requirements document, an interface …

نسخة أولية وصول مفتوح

WorldAttention: An Efficient Attention Architecture for Interactive Video World Models

Zeyu Zhang, Jinyuan Mao, Dakai An وآخرون · 2026

Leveraging the paradigm of autoregressive diffusion, text-conditioned interactive video world models aim to simulate temporally coherent environments guided by textual instructions. While enabling low-latency, long-duration generation is pivotal for embodied AI and simulation-based planning, current frameworks primaril …

نسخة أولية وصول مفتوح

Invisible in Space, Visible in Time: Motion Vision CAPTCHA against GUI Agents

Zeyu Zhang, Dingyi Rong, Zijian Chen وآخرون · 2026

Most existing visual CAPTCHAs remain spatially solvable: the required information is exposed by static appearance, local structure, and interface state. This assumption is weakened by advances in multimodal large language models (MLLMs) and Graphical User Interface (GUI) agents, which exhibit strong visual perception, …

نسخة أولية وصول مفتوح

DeltaWAM: Delta World Action Models for Bimanual Manipulation

Han Yan, Zishang Xiang, Haokai Jiang وآخرون · 2026

World-action models (WAMs) transfer visual and motion priors from pretrained video generators to robot control by jointly modeling visual dynamics and actions. Existing WAMs, however, predict dense future frames during training, repeatedly modeling largely unchanged content and coupling action-conditioned dynamics to n …

المؤلفون المشاركون