Pith. sign in

REVIEW 1 cited by

Understanding Data Importance in Machine Learning Attacks: Does Valuable Data Pose Greater Harm?

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2409.03741 v1 pith:VVZLXLNC submitted 2024-09-05 cs.CR cs.LG

classification cs.CRcs.LG
keywords datalearningmachineattacksimportancemembershipvaluableinference
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Machine learning has revolutionized numerous domains, playing a crucial role in driving advancements and enabling data-centric processes. The significance of data in training models and shaping their performance cannot be overstated. Recent research has highlighted the heterogeneous impact of individual data samples, particularly the presence of valuable data that significantly contributes to the utility and effectiveness of machine learning models. However, a critical question remains unanswered: are these valuable data samples more vulnerable to machine learning attacks? In this work, we investigate the relationship between data importance and machine learning attacks by analyzing five distinct attack types. Our findings reveal notable insights. For example, we observe that high importance data samples exhibit increased vulnerability in certain attacks, such as membership inference and model stealing. By analyzing the linkage between membership inference vulnerability and data importance, we demonstrate that sample characteristics can be integrated into membership metrics by introducing sample-specific criteria, therefore enhancing the membership inference performance. These findings emphasize the urgent need for innovative defense mechanisms that strike a balance between maximizing utility and safeguarding valuable data against potential exploitation.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. CompLeak: Deep Learning Model Compression Exacerbates Privacy Leakage

    cs.CR 2025-07 conditional novelty 5.0 of 10

    Compression of deep learning models can increase privacy leakage when multiple compressed versions are available to an attacker, and combining their outputs makes membership inference attacks much stronger.

Pith tools