← Back to all articles
arXiv cs.LGOctober 7, 2026

ATLAS-AL: Adaptive Trust-Region for Latent Adversarial Searches via Active Learning

Excerpt

arXiv:2610.07323v1 Announce Type: new Abstract: Security evaluation of learning-based systems requires more than just testing the system against a fixed collection of attacks. It requires adaptive mechanisms that can efficiently discover \textit{sets} of inputs that induce model failure. We introduce ATLAS (Adaptive Trust-Regions for Latent Adversarial Searches), which is a query-based framework that discovers adversarial input sets for black-box learning systems. ATLAS casts attack generation a