Preprint Open access
When Tools Lie: Reliability of Mathematical Agents Under Corrupted Tool Feedback
Mathematical problem solving often requires deterministic computational steps that agents delegate to tools and implicitly trust. Yet tools can fail silently, returning plausible but incorrect results. How well can agents detect and correct corrupted tool call outputs? We study this through a controlled corruption fram …