نسخة أولية وصول مفتوح
Stateless Language Agents: Scaling Long-Horizon Automated Research
Automated research systems increasingly run LLM agents over long horizons, but more inference does not by itself produce more progress: agents replay growing histories, duplicate one another's work, or stop experimenting while token consumption continues. Yet most evaluations use short budgets or benchmarks that satura …