الباحثون

Jiacheng Zhang

المنشورات 2

نسخة أولية وصول مفتوح

Adaptive Reward Routing: Dynamic Multi-Reward Optimization for Joint Audio-Video Diffusion via Forward-Process RL

Songlin Yang, Xiaotong Zhao, Jiacheng Zhang وآخرون · 2026

Multi-reward guided reinforcement learning (i.e., RL) offers a promising way to improve joint audio-video diffusion models along several complementary objectives, including modality-specific quality, cross-modal semantic alignment, and temporal synchronization. Its effectiveness, however, depends on two quantities that …

نسخة أولية وصول مفتوح

LongCat-DeepResearch Technical Report

Meituan LongCat Team, He Zhu, Yue Xu وآخرون · 2026

We present LongCat-DeepResearch, a deep research system that combines an enhanced LongCat model with a multi-agent workflow for producing comprehensive, evidence-grounded reports. The workflow separates global planning from detailed investigation and coordinates revision at the section level. Multiple planning agents f …

المؤلفون المشاركون