REVIEW 2 cited by
Attentive CutMix: An Enhanced Data Augmentation Approach for Deep Learning Based Image Classification
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Convolutional neural networks (CNN) are capable of learning robust representation with different regularization methods and activations as convolutional layers are spatially correlated. Based on this property, a large variety of regional dropout strategies have been proposed, such as Cutout, DropBlock, CutMix, etc. These methods aim to promote the network to generalize better by partially occluding the discriminative parts of objects. However, all of them perform this operation randomly, without capturing the most important region(s) within an object. In this paper, we propose Attentive CutMix, a naturally enhanced augmentation strategy based on CutMix. In each training iteration, we choose the most descriptive regions based on the intermediate attention maps from a feature extractor, which enables searching for the most discriminative parts in an image. Our proposed method is simple yet effective, easy to implement and can boost the baseline significantly. Extensive experiments on CIFAR-10/100, ImageNet datasets with various CNN architectures (in a unified setting) demonstrate the effectiveness of our proposed method, which consistently outperforms the baseline CutMix and other methods by a significant margin.
Forward citations
Cited by 2 Pith papers
-
Informed Mixing -- Improving Open Set Recognition via Attribution-based Augmentation
GradMix masks the most attribution-activated image regions during training, pushing the model to learn additional features and improving open set recognition and robustness.
-
AdaptoVision: A Multi-Resolution Image Recognition Model for Robust and Scalable Classification
AdaptoVision combines residual, depthwise, and hierarchical skip connections for image classification, but its state-of-the-art claims are contradicted by its own comparison tables and no code is provided.
Discussion (0). Continue with ORCID to comment.