الباحثون

Linfeng Zhang

المنشورات 6

نسخة أولية وصول مفتوح

Recursive Harness Self-Improvement for Frontier Reasoning Data Synthesis

Wenlong Zhang, Zhengbo Jiao, Chenxu Zhang وآخرون · 2026

Generating progressively harder reasoning problems requires synthesis procedures that adapt as the task distribution evolves. Existing task-level recursion reuses generated problems as seeds but leaves the construction harness unchanged. We present task-harness co-evolution, a framework for recursive harness self-impro …

نسخة أولية وصول مفتوح

WorkGenesis: Building the Worlds That Teach Agents to Work

Xinyu Zhu, Fenyi Liu, Yuzhu Cai وآخرون · 2026

The ability of Large Language Model (LLM) agents to complete daily and professional work is receiving increasing attention. Training such agents requires realistic work scenarios. Expert-authored occupational work is costly and slow to produce, while unconstrained synthesis often yields tasks with weak factual groundin …

نسخة أولية وصول مفتوح

Beyond Token Alignment: Event Completion for Cross-Tokenizer On-Policy Distillation

Jiacheng Liu, Jingwei Song, Qituan Zhang وآخرون · 2026

On-policy distillation (OPD) transfers knowledge between language models through teacher supervision on student-generated trajectories. With different tokenizers, a single teacher token may require multiple student tokens to generate, creating intermediate states where the event is entered but not yet completed. Existi …

نسخة أولية وصول مفتوح

Dense to MoE Adaptation for Compact Vision Language Action Policies

Muchun Niu, Shuang Chen, Yuzhou Wu وآخرون · 2026

Vision language action (VLA) policies continue to grow in parameter count, making deployment on resource-constrained robot platforms difficult. The central goal is to reduce the number of LLM-side parameters retained in the deployed policy while preserving downstream task performance. Our approach, AdaDE, adapts select …

المؤلفون المشاركون