Pith. sign in

REVIEW 1 cited by

Defending Against Physically Realizable Attacks on Image Classification

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1909.09552 v2 pith:C5PL73VW submitted 2019-09-20 cs.LG cs.AIcs.CVeess.IVstat.ML

Defending Against Physically Realizable Attacks on Image Classification

classification cs.LG cs.AIcs.CVeess.IVstat.ML
keywords attacksadversarialimageclassificationphysicallyrealizableapproachesdefending
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

We study the problem of defending deep neural network approaches for image classification from physically realizable attacks. First, we demonstrate that the two most scalable and effective methods for learning robust models, adversarial training with PGD attacks and randomized smoothing, exhibit very limited effectiveness against three of the highest profile physical attacks. Next, we propose a new abstract adversarial model, rectangular occlusion attacks, in which an adversary places a small adversarially crafted rectangle in an image, and develop two approaches for efficiently computing the resulting adversarial examples. Finally, we demonstrate that adversarial training using our new attack yields image classification models that exhibit high robustness against the physically realizable attacks we study, offering the first effective generic defense against such attacks.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Detectors Learn the Wrong Thing: Shortcut-Resistant Adversarial Training Against Physically Realizable Attacks

    cs.CV 2026-07 conditional novelty 7.0

    InsCAT adds a contrastive loss that aligns adversarially clothed people with clean people and pushes away from texture-only images, reducing texture false positives from 46.9% to 7.3% while lifting average attack AP to 82.3%.