Authors

Yanfeng Wang

Publications 2

Preprint Open access

Text-Centric Post-Training for Omni-Modal Reasoning

Improving joint audio-visual reasoning in Omni Large Language Models typically incurs substantial data construction and training costs. Our diagnostics reveal multi-hop reasoning difficulties despite correct answers to all corresponding single-hop questions and suggest partial decoupling in the local optimization of pe …

Preprint Open access

SkillCome: Group Contrast Skill Optimization with Dual Memory

Haolin Li, Feng Hong, Ang Li et al. · 2026

Skill evolution improves the capabilities of large language models by analyzing trajectories generated under a given skill and modifying the skill accordingly. Existing approaches typically generate a single trajectory per question. However, this provides insufficient optimization signals since it requires inferring ef …

Co-authors