الباحثون

Jie Cao

المنشورات 7

نسخة أولية وصول مفتوح

MMVistaReason: Toward Open-Data and Post-Training Recipes for Multimodal Reasoning

Juekai Lin, Honglin Lin, Yuqian Yuan وآخرون · 2026

Open multimodal reasoning models have benefited from large-scale reasoning supervision, yet reliable post-training remains challenging due to uneven data quality, inefficient supervision construction, imbalanced difficulty, and cross-domain interference. We introduce MMVistaReason (MVR), an open-data post-training reci …

نسخة أولية وصول مفتوح

Look Closer: Patch-wise Supervision for AI-Generated Image Detection

Zhida Zhang, Tao Wu, Siyu Liu وآخرون · 2026

How much of an image does a detector need to see? Small RGB regions can retain useful evidence of image synthesis even when they reveal little of the full scene. Motivated by single-patch detection, we study patch-wise supervision: a shared backbone classifies explicit crops, each crop receives its own loss, and patch …

نسخة أولية وصول مفتوح

GaugeVLM: Structuring Spatial Supervision with Measured Geometric Interventions

Hongbo Wang, Zihan Lin, Wenkui Yang وآخرون · 2026

Vision-language models (VLMs) can contradict themselves across views of the same spatial relation and fail to respond when that relation changes. Addressing these failures requires supervision that captures error magnitude and geometric dependencies across observations, both of which remain implicit in training on indi …

نسخة أولية وصول مفتوح

TGRL: Temperature-Grouped Reinforcement Learning for Efficient Exploration in LLMs

Zihan Lin, Xiaohan Wang, Jie Cao وآخرون · 2026

Efficient exploration often remains a central bottleneck in reinforcement learning with verifiable rewards (RLVR). Although temperature control and test-time scaling strategies can increase rollout diversity of large language models (LLMs), they either expand the sample budget at rollout time or leave the benefit of ex …

نسخة أولية وصول مفتوح

Teaching Reinforcement Learning and Humanoid Robotics to High-School Students: An Expert-Validated Curriculum Design on a Low-Cost Open Platform

Lower cost open source robots and reinforcement learning (RL) simulation tools create new opportunities for precollege students to engage with contemporary robotics. However, translating a complete research workflow, spanning mechanical assembly, electrical setup, simulation, policy learning, system identification, and …

المؤلفون المشاركون