الباحثون

Daniel Cremers

المنشورات 6

نسخة أولية وصول مفتوح

MultiFly: A Real-World Multimodal Aerial Dataset with Annotation-Efficient Label Transfer and Cross-Modal Semantic Consistency

Markus Gross, Andreas Greiner, Taehyoung Kim وآخرون · 2026

We introduce MultiFly, a real-world, low-altitude UAV dataset for semantic perception across RGB, thermal, LiDAR, and radar modalities. MultiFly provides 17,272 synchronized samples from four suburban scenes with frame-wise annotations for 15 semantic classes, together with calibration and GNSS-RTK/IMU measurements. To …

نسخة أولية وصول مفتوح

Shared Geometry As A Rosetta Stone: Cross-Modal Alignment Without Paired Data

Multimodal representations enable zero-shot classification and retrieval, but aligning independently trained models usually requires large amounts of paired data. Yet, the Platonic Representation Hypothesis suggests that models trained on different modalities may converge spontaneously toward a shared representation ge …

نسخة أولية وصول مفتوح

Lens Flare Removal and Reconstruction

The presence of lens flares in images can significantly reduce the quality of downstream application results for tasks such as 3D scene reconstruction. This is because lens flares are a property of the camera imaging system, and not a part of the underlying scene being modeled. There are previous methods that tackle th …

نسخة أولية وصول مفتوح

ORMA: Optimization-based Monocular 4D Reconstruction of Articulated Animals

Xuyi Hu, Francesco Palandra, Shangzhe Wu وآخرون · 2026

Recovering articulated 4D representations of animals from monocular videos remains challenging due to the large diversity of quadruped morphologies and lack of animal 4D supervision data. Existing learning-based reconstruction methods operate on individual images and rely on synthetic or model-fitted 3D supervision, wh …

المؤلفون المشاركون