نسخة أولية وصول مفتوح
Maintaining Benchmarks Against Increasingly Capable Agents: Detection and Remediation of Unearned Passes
Agentic benchmarks guide model selection and training. Yet an agent can pass a task without demonstrating the intended capability. Such outcomes constitute unearned passes; their proportion among all passes defines the integrity gap. As agents improve, benchmark surfaces that once seemed harmless can become exploitable …