الباحثون

Xiangbo Gao

المنشورات 3

نسخة أولية وصول مفتوح

ViTeX-Bench: Benchmarking High-Fidelity Video Scene Text Editing

Xinghao Chen, Xiangbo Gao, Jiongze Yu وآخرون · 2026

Recent video generation is increasingly realistic and controllable, yet video editing remains less developed, particularly for precise local edits that must preserve the original scene dynamics. Video scene text editing replaces text on scene surfaces, such as storefront signs, whiteboards, and product labels, while pr …

نسخة أولية وصول مفتوح

Joint and Cross-Modal Video-Audio Generation and Editing: A Unified Formulation and Design Taxonomy

Video and audio are perceived together, yet most generative models treat them in isolation. We examine methods that model the two modalities jointly, generate one from the other, or edit them in a coupled manner, organized around a single question: how is the output kept coherent across modalities in time and semantics …

نسخة أولية وصول مفتوح

AURORA: A Natural Language-Driven Agentic Framework for Understanding, Reasoning, and Orchestrating Reliable Air-Ground Co-Simulation

Keshu Wu, Hao Zhang, Rui Gan وآخرون · 2026

Air-ground transportation research increasingly relies on co-simulation, yet constructing scenarios remains labor-intensive and difficult to validate. More importantly, a generated scenario may execute successfully while failing to realize the spatial, temporal, communication, or behavioral relationships requested by t …

المؤلفون المشاركون