Pith. sign in

REVIEW 2 cited by

Robustness through Data Augmentation Loss Consistency

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2110.11205 v3 pith:I7MBTB6Y submitted 2021-10-21 cs.LG cs.AIcs.CLcs.CV

Robustness through Data Augmentation Loss Consistency

classification cs.LG cs.AIcs.CLcs.CV
keywords dataaugmentationrobustdairconsistencycovariantregularizationaugmented
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

While deep learning through empirical risk minimization (ERM) has succeeded at achieving human-level performance at a variety of complex tasks, ERM is not robust to distribution shifts or adversarial attacks. Synthetic data augmentation followed by empirical risk minimization (DA-ERM) is a simple and widely used solution to improve robustness in ERM. In addition, consistency regularization can be applied to further improve the robustness of the model by forcing the representation of the original sample and the augmented one to be similar. However, existing consistency regularization methods are not applicable to covariant data augmentation, where the label in the augmented sample is dependent on the augmentation function. For example, dialog state covaries with named entity when we augment data with a new named entity. In this paper, we propose data augmented loss invariant regularization (DAIR), a simple form of consistency regularization that is applied directly at the loss level rather than intermediate features, making it widely applicable to both invariant and covariant data augmentation regardless of network architecture, problem setup, and task. We apply DAIR to real-world learning problems involving covariant data augmentation: robust neural task-oriented dialog state tracking and robust visual question answering. We also apply DAIR to tasks involving invariant data augmentation: robust regression, robust classification against adversarial attacks, and robust ImageNet classification under distribution shift. Our experiments show that DAIR consistently outperforms ERM and DA-ERM with little marginal computational cost and sets new state-of-the-art results in several benchmarks involving covariant data augmentation. Our code of all experiments is available at: https://github.com/optimization-for-data-driven-science/DAIR.git

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Unsupervised learning for the systematic identification of nondispersive wave packets in driven helium

    quant-ph 2026-05 unverdicted novelty 5.0

    Unsupervised CNN embedding and clustering of Floquet states recovers known nondispersive wave packet regimes in driven helium without labels.

  2. Geometrically Constrained and Token-Based Probabilistic Spatial Transformers

    cs.CV 2025-09 conditional novelty 5.0

    A probabilistic, component-wise spatial transformer using a shared frozen tokenizer and augmentation-guided alignment loss improves geometric robustness in fine-grained moth classification.