Adversarial Training with Orthogonal Regularization
2020
The study of how learning systems behave under deliberately chosen perturbations intended to cause errors or expose vulnerabilities.
It includes threat-model specification, adversarial evaluation, and methods for training systems to resist the specified attacks.