الباحثون

Belinda Zeng

المنشورات 2

نسخة أولية وصول مفتوح

Native Action-Prior Learning from Videos for World Action Models

Zhaochong An, Fei Zhang, Menglin Jia وآخرون · 2026

World action models integrate future visual dynamics with robot action prediction, but their scalability remains limited by the need for action-annotated robot trajectories. Observation-only videos contain rich evidence about interaction dynamics, but existing approaches typically use them either to pretrain visual rep …

نسخة أولية وصول مفتوح

Unifying Video Tasks via Spatiotemporal Analogy

Adapting video models to new tasks typically requires dedicated data curation and fine-tuning. While visual analogy provides a training-free alternative by specifying tasks in-context, it remains restricted to the image domain. To explore whether analogy-based methods can unify diverse video tasks and generalize to out …

المؤلفون المشاركون