Preprint Open access
ATLAS-AL: Adaptive Trust-Region for Latent Adversarial Searches via Active Learning
Security evaluation of learning-based systems requires more than just testing the system against a fixed collection of attacks. It requires adaptive mechanisms that can efficiently discover \textit{sets} of inputs that induce model failure. We introduce ATLAS (Adaptive Trust-Regions for Latent Adversarial Searches), wh …