الباحثون

Yuchen Yan

المنشورات 4

نسخة أولية وصول مفتوح

Distilling Routed 3D Privilege for Spatial Reasoning in Vision-Language Models

Hongxing Li, Yixin Li, Dingming Li وآخرون · 2026

Spatial reasoning remains a persistent weakness of vision-language models (VLMs), because RGB inputs do not directly provide geometric evidence. Existing remedies either inject 3D into the model at inference, paying architecture and latency costs, or train with outcome rewards that supervise only the final answer. Spat …

نسخة أولية وصول مفتوح

Video2World: Benchmarking Coding Agents for Interactive World Modeling from Embodied Videos

Jinzhou Tang, Zijun Zhang, Jing Yang وآخرون · 2026

Building interactive simulators from real-world observations is a promising way to scale embodied data, but current pipelines still rely heavily on manual environment construction and calibration. We study whether frontier foundation models and coding agents can automate this process end to end. We formulate \emph{auto …

نسخة أولية وصول مفتوح

IterSynth: Rethinking Deep Search Agents via Role-Decoupled Iterative Synthesis

Xingyu Wu, Yuchen Yan, Zhengxi Lu وآخرون · 2026

Deep search requires LLM agents to decompose complex queries, search for evidence, and synthesize grounded answers, yet existing ReAct-style agents suffer from two limitations: role coupling, where one policy must handle planning, evidence use, and synthesis; and context accumulation, where growing search histories int …

المؤلفون المشاركون