Pith. sign in

REVIEW 2 cited by

Feature Importance Guided Attack: A Model Agnostic Adversarial Attack

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2106.14815 v3 pith:PVD6PZFU submitted 2021-06-28 cs.LG cs.AIcs.CR

classification cs.LGcs.AIcs.CR
keywords featurefigaspaceattackdatasetsphishingproblemtabular
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Research in adversarial learning has primarily focused on homogeneous unstructured datasets, which often map into the problem space naturally. Inverting a feature space attack on heterogeneous datasets into the problem space is much more challenging, particularly the task of finding the perturbation to perform. This work presents a formal search strategy: the `Feature Importance Guided Attack' (FIGA), which finds perturbations in the feature space of heterogeneous tabular datasets to produce evasion attacks. We first demonstrate FIGA in the feature space and then in the problem space. FIGA assumes no prior knowledge of the defending model's learning algorithm and does not require any gradient information. FIGA assumes knowledge of the feature representation and the mean feature values of defending model's dataset. FIGA leverages feature importance rankings by perturbing the most important features of the input in the direction of the target class. While FIGA is conceptually similar to other work which uses feature selection processes (e.g., mimicry attacks), we formalize an attack algorithm with three tunable parameters and investigate the strength of FIGA on tabular datasets. We demonstrate the effectiveness of FIGA by evading phishing detection models trained on four different tabular phishing datasets and one financial dataset with an average success rate of 94%. We extend FIGA to the phishing problem space by limiting the possible perturbations to be valid and feasible in the phishing domain. We generate valid adversarial phishing sites that are visually identical to their unperturbed counterpart and use them to attack six tabular ML models achieving a 13.05% average success rate.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. FitCF: A Framework for Automatic Feature Importance-guided Counterfactual Example Generation

    cs.CL 2025-01 conditional novelty 6.0 of 10

    ZeroCF and FitCF generate label-flipping text counterfactuals from BERT feature attributions, and FitCF outperforms Polyjuice, BAE, and FIZLE on AG News and SST2.

  2. Insights on Adversarial Attacks for Tabular Machine Learning via a Systematic Literature Review

    cs.LG 2025-06 conditional novelty 5.0 of 10

    A systematic review of 53 papers on adversarial attacks for tabular machine learning finds the field fragmented, with efficacy over-emphasized and transferability, plausibility, and semantic preservation under-addressed.

Pith tools