الملخص
تمت ترجمة أجزاء من هذه الصفحة آلياً وقد تحتوي على أخطاء.
Harness optimization provides a practical setting for recursive self-improvement (RSI), where agent-generated modifications inform subsequent changes through execution feedback. Recent work such as Meta-Harness implements this process through iterative code generation and evaluation, but retains a fixed development set and proposal policy. These constraints channel evolution along a single search trajectory, increasing the risk of converging to a local optimum. We make the improvement process itself adaptive by organizing search into branches with evolving development subsets and proposal policies. Each branch retains development cases solved by more of its leading harnesses than by those of other branches, drops cases solved by every leading harness across all branches, and revises its proposal policy using its own search history. To deploy the resulting complementary harnesses, we propose a router to select one development-selected branch head for each new input before execution. Across mathematical reasoning and agentic coding benchmarks, our system achieves relative improvements over Meta-Harness of 34.8% on Olympiad-level mathematical reasoning, 11.6% on Terminal-Bench 2.0, and 3.8% on SWE-bench Lite, with harness selection and router configuration based solely on development data. These results show that evolving branch objectives and proposal policies can yield complementary harnesses whose strengths a router combines without access to test outcomes.
الكلمات المفتاحية
الموضوع
بيانات النشر
- المجلة
- غير متاح
- وصول مفتوح
- وصول مفتوح أخضر
اقتبس هذه المقالة
APA 7
Dong, H., Zhou, Y., Lin, Z., Wu, Y., Peng, B., Wang, M., Fan, X., Zhang, L., & Zhao, Z. (2026). خليط من الفروع ذاتية التحسين لتحسين أحزمة الوكلاء. https://omanscience.com/ar/articles/mixture-of-self-improving-branches-for-agent-harness-optimization
MLA 9
Dong, Haoyu, et al. "خليط من الفروع ذاتية التحسين لتحسين أحزمة الوكلاء." https://omanscience.com/ar/articles/mixture-of-self-improving-branches-for-agent-harness-optimization.
شيكاغو (المؤلف–التاريخ)
Dong, Haoyu, Yuhang Zhou, Zihao Lin, Yifan Wu, Bo Peng, Mingyi Wang, Xiangjun Fan, Lizhu Zhang, and Zhuokai Zhao. 2026. "خليط من الفروع ذاتية التحسين لتحسين أحزمة الوكلاء." https://omanscience.com/ar/articles/mixture-of-self-improving-branches-for-agent-harness-optimization.
هارفارد
Dong, H., Zhou, Y., Lin, Z., Wu, Y., Peng, B., Wang, M., Fan, X., Zhang, L. and Zhao, Z. (2026) 'خليط من الفروع ذاتية التحسين لتحسين أحزمة الوكلاء', Available at: https://omanscience.com/ar/articles/mixture-of-self-improving-branches-for-agent-harness-optimization.
فانكوفر
Dong H, Zhou Y, Lin Z, Wu Y, Peng B, Wang M, et al. خليط من الفروع ذاتية التحسين لتحسين أحزمة الوكلاء. https://omanscience.com/ar/articles/mixture-of-self-improving-branches-for-agent-harness-optimization
IEEE
H. Dong, Y. Zhou, Z. Lin, Y. Wu, B. Peng, M. Wang, X. Fan, L. Zhang, and Z. Zhao, "خليط من الفروع ذاتية التحسين لتحسين أحزمة الوكلاء," https://omanscience.com/ar/articles/mixture-of-self-improving-branches-for-agent-harness-optimization.