الباحثون

Yang Shi

المنشورات 5

نسخة أولية وصول مفتوح

Dynamics-Aware Adaptive Corridors with Feasibility-Perturbed Trust-Region SQP for Certified Nonholonomic Motion Planning

Yang Shi · 2026

Optimisation-based parking planners usually impose collision constraints only at the time samples, so a vehicle corner can cut an obstacle between samples, and no executable trajectory exists until the solver converges. We present a planner for car-like vehicles with reverse gear in which every iterate of the optimisat …

نسخة أولية وصول مفتوح

PortraitAes: Intent-Conditioned Structured Portrait Aesthetics Assessment

Junzhou Xie, Haozhong Xiong, Xunyun Tian وآخرون · 2026

Portrait aesthetic assessment assigns comparable scores according to how effectively human-centered images fulfill their photographic intent. These scores support data filtering, candidate selection, and preference modeling in image-generation pipelines. Existing methods typically predict a single aesthetic score or us …

نسخة أولية وصول مفتوح

Answer with Evidence: Consistency-Aware Grounded Visual Question Answering for Roadside Traffic Scenes

Runwei Guan, Rongsheng Hu, Shangshu Chen وآخرون · 2026

Roadside traffic reasoning requires every free-form textual claim to be backed by visual evidence. Existing grounded multimodal large language models (MLLMs) frequently exhibit say-point mismatch, in which the textual answer contradicts the bounding boxes the model localizes. Evaluation metrics that score answers and b …

نسخة أولية وصول مفتوح

Think Before You Score: Thinking Reward Model for Visual Generation

Xuehai Bai, Zhenchen Tang, Yang Shi وآخرون · 2026

Visual reward models are essential for evaluating and improving visual generation models, yet existing approaches typically map task conditions and candidate outputs directly to scalar rewards, leaving implicit what should be evaluated for each individual case. We introduce Think Before You Score, a paradigm that expli …

نسخة أولية وصول مفتوح

HiRAE: Hierarchical Representation Autoencoding with Residual Budgets

Xuanyu Zhu, Yan Bai, Yang Shi وآخرون · 2026

Pretrained visual representations support image generation, but may not fully preserve the fine-grained details needed for faithful reconstruction. Meanwhile, intermediate encoder layers contain complementary visual details, but learning to fuse them for reconstruction can produce a latent distribution that is difficul …

المؤلفون المشاركون