REVIEW 6 cited by
A Survey on Transferability of Adversarial Examples across Deep Neural Networks
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The emergence of Deep Neural Networks (DNNs) has revolutionized various domains by enabling the resolution of complex tasks spanning image recognition, natural language processing, and scientific problem-solving. However, this progress has also brought to light a concerning vulnerability: adversarial examples. These crafted inputs, imperceptible to humans, can manipulate machine learning models into making erroneous predictions, raising concerns for safety-critical applications. An intriguing property of this phenomenon is the transferability of adversarial examples, where perturbations crafted for one model can deceive another, often with a different architecture. This intriguing property enables black-box attacks which circumvents the need for detailed knowledge of the target model. This survey explores the landscape of the adversarial transferability of adversarial examples. We categorize existing methodologies to enhance adversarial transferability and discuss the fundamental principles guiding each approach. While the predominant body of research primarily concentrates on image classification, we also extend our discussion to encompass other vision tasks and beyond. Challenges and opportunities are discussed, highlighting the importance of fortifying DNNs against adversarial vulnerabilities in an evolving landscape.
Forward citations
Cited by 6 Pith papers
-
3D Gaussian Splatting Driven Multi-View Robust Physical Adversarial Camouflage Generation
PGA uses 3D Gaussian Splatting to generate physical adversarial camouflage from a few images, improving multi-view attack robustness on vehicle detectors.
-
I Stolenly Swear That I Am Up to (No) Good: Design and Evaluation of Model Stealing Attacks
A systematization of 47 model stealing papers that introduces a threat model, a comparison framework, and evaluation best practices for substitute-model attacks.
-
Exploiting Edge Features for Transferable Adversarial Attacks in Distributed Machine Learning
Intercepting intermediate features in split neural network inference lets black-box attackers build surrogate models whose adversarial examples transfer to the target far more often, e.g., 96% versus 61% success in on...
-
Light as Deception: GPT-driven Natural Relighting Against Vision-Language Pre-training Models
LightD creates natural adversarial relighting images with GPT-selected lighting parameters and gradient optimization, outperforming prior non-suspicious attacks on vision-language models.
-
DUMB and DUMBer: Is Adversarial Training Worth It in the Real World?
Across 240 model configurations and 13 attacks, adaptive and curriculum adversarial training give the largest robustness gains, but 20.53% of evaluations show negative gains, mostly under mismatched source-target mode...
-
Understanding Knowledge Transferability for Transfer Learning: A Survey
A survey that classifies transferability metrics by knowledge modality (dataset vs. model) and granularity (task vs. instance), with a theoretical primer and applications to eight learning paradigms.
Discussion (0). Sign in to comment.