الباحثون

Qianqian Wang

المنشورات 3

نسخة أولية وصول مفتوح

MoSE3: Learning World-Space SE(3) at Every Pixel

Jiahuan Cheng, Zhiyi Li, Tian Xia وآخرون · 2026

Dense 3D point tracking has been a prominent paradigm for modeling motion in dynamic scenes, but a point track is just a 3-DoF translation curve per pixel: it captures where pixels go, not the rotation of the underlying part, nor which pixels move together as one body. We propose MoSE3, the first feed-forward model tha …

نسخة أولية وصول مفتوح

MosaiChunk: Compositing Spatio-Temporal Memory for Autoregressive Video Generation

Long-horizon autoregressive video generation is limited by a finite context window. When an object or scene falls out of context, its fine-grained visual details may be lost and difficult to recover upon reappearance. To retain access to such visual details, we introduce MosaiChunk, a spatio-temporal memory mechanism t …

المؤلفون المشاركون