الباحثون

Xuezhe Ma

المنشورات 5

نسخة أولية وصول مفتوح

SynCo: Data Synthesis Co-Training for Self-Evolving LLMs via Multi-Agent Reinforcement Learning

Wei Yang, Shawn Li, Yuehan Qin وآخرون · 2026

Self-evolving LLM agents promise to improve autonomously through continual interaction and learning, reducing their dependence on manually curated supervision. Realizing this promise requires not only updating the agent, but also evolving its training experience as its capabilities change. However, most existing pipeli …

نسخة أولية وصول مفتوح

Visual Abstention in Unified Multimodal Models

Chufan Shi, Cheng Yang, Tiannuo Yang وآخرون · 2026

Unified multimodal models (UMMs) integrate understanding and generation, yet their generative behavior is rarely governed by what they understand about the task. We formalize visual abstention: when a requested visual transformation is impossible under the task's rules, the model should recognize that no valid solution …

نسخة أولية وصول مفتوح

Recursive Self-Improvement in Unified Multimodal Models

Huijuan Wang, Chufan Shi, Cheng Yang وآخرون · 2026

Unified multimodal models (UMMs) understand and generate both text and images, which lets a model produce its own training data. Existing self-improvement in UMMs keeps supervision on the visual side, where image understanding judges image generation. We propose recursive cross-capability self-improvement (RSI), a trai …

نسخة أولية وصول مفتوح

InfoEdit: Probing Global Layout Reasoning in Infographic Editing

Cheng Yang, Chufan Shi, Huijuan Wang وآخرون · 2026

Multimodal foundation models edit natural photographs at production quality, yet the same models struggle with structured visual content such as infographics. Unlike photographs, infographics encode information through logical relations; editing one element often requires surrounding elements to be adapted. We refer to …

المؤلفون المشاركون