نسخة أولية وصول مفتوح
Long-Horizon Scaling: How Model Capabilities Shape the Returns to Computation
Long-horizon agents improve solutions through sustained interaction, execution, and task feedback. Scaling studies relate performance to resources and capabilities, yet how existing capabilities shape returns to extended interaction remains less understood. To address this gap, we analyze AutoLab and EdgeBench, two lon …