الباحثون

Xiangyang Xue

المنشورات 4

نسخة أولية وصول مفتوح

OmniCam: Omni-Camera Trajectory Generation via Geometry-Grounded Pose Token Learning

Zhenyang Liu, Chenjie Cao, Yisu Zhang وآخرون · 2026

Camera trajectories control viewpoint changes in video generation, scene reconstruction, and robotic perception. Generating them from language requires both scene geometry and target-aware framing. We introduce OmniCam, an autoregressive model that generates camera pose sequences from a single panorama and textual traj …

نسخة أولية وصول مفتوح

Anchor and Adapt: Asymmetric Prompt Adaptation for Few-Shot Industrial Anomaly Detection

Mengyang Zhao, Teng Fu, Haiyang Yu وآخرون · 2026

In few-shot industrial anomaly detection, the few normal target images provide no direct defect supervision, making anomaly prompts difficult to learn from these samples alone. Some vision-language methods therefore use manually specified descriptions to supply explicit anomaly semantics. However, constructing these de …

نسخة أولية وصول مفتوح

CAR-VLA: Complexity-Aware and Risk-Adaptive Reasoning for Autonomous Driving

Xiaolei Chen, Zhuolin He, Yuxuan Liang وآخرون · 2026

Existing adaptive reasoning methods for driving Vision-Language-Action (VLA) models primarily focus on whether to reason, overlooking how reasoning should differ across driving situations. Our key insight is that while scene complexity informs reasoning depth, dynamic risk is equally critical for deciding how to reason …

المؤلفون المشاركون