الباحثون

Luc Van Gool

المنشورات 9

نسخة أولية وصول مفتوح

iSEE: Object Permanence Through Self-Supervision

Pramish Paudel, Ajad Chhatkuli, Luc Van Gool وآخرون · 2026

Object permanence, keeping track of an object's identity and position while it is occluded, is central to video representations that track, predict and plan. Trackers that achieve it learn from boxes, track identities and visibility labels. On the other hand, self-supervised object-centric methods discover objects with …

نسخة أولية وصول مفتوح

Ego-Forge: Text and Geometric-Attention Free Exo-to-Egocentric Video Generation

Exo-to-egocentric video generation aims to synthesize what a person sees from their own viewpoint given third-person footage and a target head trajectory. The task requires transferring appearance and semantics across large viewpoint changes while hallucinating content never observed by the exocentric camera. Existing …

نسخة أولية وصول مفتوح

AHMAD: Adaptive Hybrid Multi-task Vision Learning with Assisted Distillation for Keypoint Detection

Generalist multitasking vision models aim to unify multiple vision tasks within a single framework, enabling more efficient and versatile learning. However, handling diverse vision tasks -- spanning dense and sparse predictions -- remains challenging due to their inherently varying output structures. In this paper, we …

نسخة أولية وصول مفتوح

GraphWrit3R: End-to-End 3D Scene Graph Writing

3D scene graphs provide a structured representation of complex environments by encoding objects, their semantic attributes, and the spatial and functional relationships between them. Current approaches for 3D scene graph generation suffer from several fundamental limitations. They rely on complex multi-stage pipelines …

نسخة أولية وصول مفتوح

LensDesigner: A Self-Improving Agent for Optical Lens Design

Lei Sun, Haoran Liang, Dannong Xu وآخرون · 2026

Optical lens design is a complex, non-convex optimization challenge that relies heavily on human experience and intuition. Existing optimized-based automatic lens design methods struggle to navigate this vast parameter space without meticulous manual tuning. In this paper, we present LensDesigner, an autonomous agent f …

نسخة أولية وصول مفتوح

CoRef-GS: Cooperative Referring Gaussian Splatting for Multi-Agent Scene Understanding

Zhikun Zhou, Kunyu Peng, Runyi Yang وآخرون · 2026

Referring scene understanding for embodied robots requires grounding object- and relation-centric language queries from a designated viewpoint. While a local semantic Gaussian map can support such grounding within one agent's observations, cooperative settings require this ability to remain effective after independentl …

المؤلفون المشاركون