الباحثون

Yunbei Zhang

المنشورات 8

نسخة أولية وصول مفتوح

Safety Must Survive Self-Improvement: Why Failures Persist and How Agents Recover

Yunbei Zhang, Janet Wang, Saiyue Lyu وآخرون · 2026

Recursive self-improvement (RSI) allows agents to carry useful changes across generations. Maintaining safety across these generations involves both preventing unsafe behavior from persisting and enabling recovery when failures occur. We study these challenges through a controlled testbed of stateful authorization task …

نسخة أولية وصول مفتوح

DeepJEPA: Scaling World Models from Within

Zijian Jin, Yunbei Zhang, Yuanzhe Liu وآخرون · 2026

World-model planners typically scale outward by rolling farther, sampling more trajectories, or optimizing longer, while assigning the same computation to every imagined transition. We show that making every transition uniformly deeper wastes computation and can degrade planning because useful refinement is concentrate …

نسخة أولية وصول مفتوح

Deny Without Disabling: Authorization-Paired Evaluation and Control for Multi-Agent Systems

Yunbei Zhang, Saiyue Lyu, Janet Wang وآخرون · 2026

Multi-agent systems derive their capabilities from sharing evidence, delegating tasks, and combining information across agents. The same process creates a safety problem: contributions that are admissible in isolation can jointly enable a prohibited use. Blocking every sensitive action avoids disclosure but defeats the …

نسخة أولية وصول مفتوح

LIBERO-MAX: Do Robot Policies Adapt When the World Changes?

Yunbei Zhang, Zijian Jin, Yuanzhe Liu وآخرون · 2026

Robots must often continue a task after a target moves, the viewpoint shifts, or an obstacle appears, even though their earlier observations and committed actions reflect the previous scene. Many simulation robustness benchmarks fix external conditions at reset, leaving this temporal challenge underexamined. We introdu …

نسخة أولية وصول مفتوح

How Medical VLMs Underutilize Their Vision Encoders: A Dermatology Perspective

Janet Wang, Yunbei Zhang, Xiao Wang وآخرون · 2026

Medical Vision-Language Models (VLMs) show significant promise for clinical image understanding, offering accurate diagnosis with interpretable reasoning. However, a critical performance gap exists between their strong vision encoders and the full multimodal model: in dermatology, the MedSigLIP encoder outperforms MedG …

نسخة أولية وصول مفتوح

Simple Agentic Memory for Generalist Robot Policies

Yuyou Zhang, Yunbei Zhang, Miao Li وآخرون · 2026

Visual-memory systems commonly retain or compress past observations. Robot control additionally requires interaction-derived state that no individual frame may explicitly represent, such as persistent identity relations, accumulated progress, or ordered procedures. We introduce Simple Agentic Robot Memory (SimpleARM), …

نسخة أولية وصول مفتوح

FORGE: Form-Optimal Routing of Grounded Evidence for Frozen LLM Agents

Xi Xiao, Yunbei Zhang, Chen Liu وآخرون · 2026

In agentic AI systems, frozen foundation models are increasingly deployed as closed-weight API endpoints, making downstream adaptation possible only through the inputs and inference procedures surrounding the model. As a result, for each input query, two coupled decisions largely determine both answer quality and token …

المؤلفون المشاركون