الباحثون

Zhuowen Tu

المنشورات 3

نسخة أولية وصول مفتوح

OverLay++: Dense-Overlap Layout-to-Image Generation Dataset

Layout-to-Image generation has made substantial progress in spatial and object-level control. However, existing methods still struggle with complex scenes containing many overlapping and interacting objects. We argue that training data is a particular bottleneck: existing datasets lack examples with dense, complex obje …

نسخة أولية وصول مفتوح

LIFT: Layout-In-Future Video Generation under Large Viewpoint Change via On-Policy Self-Distillation

Shengxiang Ji, Boyang Wang, Haiyang Xu وآخرون · 2026

We introduce LIFT, a unified image-to-video generation framework that complements camera control with Layout-In-FuTure control, enabling users to specify what should appear in a future view and where it should appear. This addresses a practical need in controllable video generation: given an initial image, users often …

المؤلفون المشاركون