الباحثون

Dong Chen

المنشورات 3

نسخة أولية وصول مفتوح

VibeEdit: Image Editing with Canvas Instructions

Jinjing Zhao, Fangyun Wei, Yitong Wang وآخرون · 2026

In text-guided image editing, describing the desired change is often straightforward, but identifying the intended object or region can be cumbersome, especially when several objects look alike. We introduce a new image editing interface that lets users place spatial marks and optional short notes directly on the image …

نسخة أولية وصول مفتوح

From Prompting to Composing: A Spatial Canvas Interface for Poster Generation

Yitong Wang, Fangyun Wei, Jinjing Zhao وآخرون · 2026

Text prompting is an indirect interface for poster generation, requiring users to encode inherently two-dimensional composition intent into a one-dimensional sequence of words. We introduce a Spatial Canvas Interface that enables users to directly compose generation intent in space through four complementary binding ty …

نسخة أولية وصول مفتوح

VidAct: Learning Manipulation from In-the-Wild Videos with Object-Centric 3D Awareness

Hang Li, Mingxin Zhang, Zihan Wu وآخرون · 2026

Video demonstrations offer a scalable alternative to costly robot data for learning manipulation, yet existing reconstruction-based approaches often rely on constrained camera viewpoints or human-to-robot retargeting, while the reconstructed trajectories are difficult to adapt to new objects configurations without dist …

المؤلفون المشاركون