نسخة أولية وصول مفتوح
VibeMemBench: Evaluating Memory Systems for Coding Agents on Real Repository Coding Tasks
Coding agents operate on real repository coding tasks, and persistent memory systems promise to reuse experience across tasks. Yet existing evaluations do not show whether those systems improve executable repository work. Repository benchmarks test code changes but do not isolate memory, while memory benchmarks score r …