الباحثون

Yiwei Li

المنشورات 2

نسخة أولية وصول مفتوح

From Pareto to Preference: Personalized Test-Time Scaling via Amortized Agentic Policy Discovery

Xinglin Wang, Zishen Liu, Tong Zheng وآخرون · 2026

Test-time scaling (TTS) improves the reasoning capabilities of large language models by allocating additional inference computation. Existing approaches to improving TTS efficiency largely optimize accuracy against one resource dimension at a time, advancing either the accuracy--cost or accuracy--latency Pareto frontie …

نسخة أولية وصول مفتوح

Long-Horizon Scaling: How Model Capabilities Shape the Returns to Computation

Haoyu Zheng, Zhengyu Chen, Huaisheng Zhu وآخرون · 2026

Long-horizon agents improve solutions through sustained interaction, execution, and task feedback. Scaling studies relate performance to resources and capabilities, yet how existing capabilities shape returns to extended interaction remains less understood. To address this gap, we analyze AutoLab and EdgeBench, two lon …

المؤلفون المشاركون