Pith. sign in

REVIEW 1 cited by

Provable Benefit of Cutout and CutMix for Feature Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2410.23672 v1 pith:FXIVIKNQ submitted 2024-10-31 cs.LG cs.AIstat.ML

classification cs.LGcs.AIstat.ML
keywords trainingcutmixcutoutfeaturesaugmentationlearnanalysiscannot
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Patch-level data augmentation techniques such as Cutout and CutMix have demonstrated significant efficacy in enhancing the performance of vision tasks. However, a comprehensive theoretical understanding of these methods remains elusive. In this paper, we study two-layer neural networks trained using three distinct methods: vanilla training without augmentation, Cutout training, and CutMix training. Our analysis focuses on a feature-noise data model, which consists of several label-dependent features of varying rarity and label-independent noises of differing strengths. Our theorems demonstrate that Cutout training can learn low-frequency features that vanilla training cannot, while CutMix training can learn even rarer features that Cutout cannot capture. From this, we establish that CutMix yields the highest test accuracy among the three. Our novel analysis reveals that CutMix training makes the network learn all features and noise vectors "evenly" regardless of the rarity and strength, which provides an interesting insight into understanding patch-level augmentation.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Towards Understanding Why Data Augmentation Improves Generalization

    cs.CV 2025-02 conditional novelty 6.0 of 10

    Data augmentation improves generalization through two mechanisms, partial semantic feature removal and feature mixing, which respectively promote diverse and robust feature learning.

Pith tools