الباحثون

Yujin Tang

المنشورات 3

نسخة أولية وصول مفتوح

Recursive Self-Improvement through Multi-Agent Self-Supervision

Hyunin Lee, Jinglue Xu, Jeffrey Seely وآخرون · 2026

Recursive self-improvement (RSI) of a model on non-verifiable tasks, such as open-ended research, faces a supervision bottleneck when its outputs exceed what even human experts can reliably assess, leaving the model itself (optimizee) as the best available optimizer and evaluator. However, a single model instance strug …

نسخة أولية وصول مفتوح

Learning and Transferring Closed-Loop Robot Software

Closed-loop robot policies require observation processing, state management, and situation-dependent branching, making them costly to design and tune manually. Although coding agents increasingly support control-code generation and optimization, it remains unclear whether implementations improved on source tasks also s …

المؤلفون المشاركون