Pith. sign in

REVIEW 2 cited by

If Influence Functions are the Answer, Then What is the Question?

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2209.05364 v1 pith:N5P3AFWA submitted 2022-09-12 cs.LG stat.ML

classification cs.LGstat.ML
keywords influencefunctionfunctionsanswerestimatesfactorsleave-one-outnetworks
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Influence functions efficiently estimate the effect of removing a single training data point on a model's learned parameters. While influence estimates align well with leave-one-out retraining for linear models, recent works have shown this alignment is often poor in neural networks. In this work, we investigate the specific factors that cause this discrepancy by decomposing it into five separate terms. We study the contributions of each term on a variety of architectures and datasets and how they vary with factors such as network width and training time. While practical influence function estimates may be a poor match to leave-one-out retraining for nonlinear networks, we show they are often a good approximation to a different object we term the proximal Bregman response function (PBRF). Since the PBRF can still be used to answer many of the questions motivating influence functions, such as identifying influential or mislabeled examples, our results suggest that current algorithms for influence function estimation give more informative results than previous error analyses would suggest.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 12 citations worldwide. Full citation record

  1. MAGIC: Near-Optimal Data Attribution for Deep Learning

    cs.LG 2025-04 conditional novelty 6.0 of 10

    MAGIC computes the exact influence function for smooth, deterministic deep learning training runs and achieves near-perfect linear datamodeling scores on CIFAR-10, GPT-2, and Gemma-2B, far outperforming TRAK and EK-FAC.

  2. Precision Profile Pollution Attack on Sequential Recommenders via Influence Function

    cs.IR 2024-12 reject novelty 5.0 of 10

    INFAttack uses influence functions to greedily pick injected items for profile pollution, reporting better target-item promotion than gradient and similarity based attacks on five datasets.

Pith tools