Pith. sign in

REVIEW 1 cited by

Revisiting adversarial training for the worst-performing class

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2302.08872 v1 pith:PQVPZYXN submitted 2023-02-17 cs.LG

classification cs.LG
keywords classtrainingworstworst-performingaccuracyadversarialcifar10classes
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Despite progress in adversarial training (AT), there is a substantial gap between the top-performing and worst-performing classes in many datasets. For example, on CIFAR10, the accuracies for the best and worst classes are 74% and 23%, respectively. We argue that this gap can be reduced by explicitly optimizing for the worst-performing class, resulting in a min-max-max optimization formulation. Our method, called class focused online learning (CFOL), includes high probability convergence guarantees for the worst class loss and can be easily integrated into existing training setups with minimal computational overhead. We demonstrate an improvement to 32% in the worst class accuracy on CIFAR10, and we observe consistent behavior across CIFAR100 and STL10. Our study highlights the importance of moving beyond average accuracy, which is particularly important in safety-critical applications.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Towards Fair Class-wise Robustness: Class Optimal Distribution Adversarial Training

    cs.LG 2025-01 conditional novelty 3.0 of 10

    A chi-squared distributionally robust optimization reweighting scheme improves worst-class robustness in adversarially trained image classifiers.

Pith tools