Pith. sign in

REVIEW

Improving Adversarial Robustness by Putting More Regularizations on Less Robust Samples

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2206.03353 v4 pith:XNBV3YIM submitted 2022-06-07 stat.ML cs.LG

classification stat.MLcs.LG
keywords adversarialalgorithmattacksrobustnessaccuracyalgorithmsdataexisting
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Adversarial training, which is to enhance robustness against adversarial attacks, has received much attention because it is easy to generate human-imperceptible perturbations of data to deceive a given deep neural network. In this paper, we propose a new adversarial training algorithm that is theoretically well motivated and empirically superior to other existing algorithms. A novel feature of the proposed algorithm is to apply more regularization to data vulnerable to adversarial attacks than other existing regularization algorithms do. Theoretically, we show that our algorithm can be understood as an algorithm of minimizing the regularized empirical risk motivated from a newly derived upper bound of the robust risk. Numerical experiments illustrate that our proposed algorithm improves the generalization (accuracy on examples) and robustness (accuracy on adversarial attacks) simultaneously to achieve the state-of-the-art performance.

Discussion (0). Continue with ORCID to comment.

Pith tools