Pith. sign in

Paper Citation Record · LEDGER

Datamodels: Predicting Predictions from Training Data

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2202.00622.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2202.00622 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:09:34.826920Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

18
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4f4febd9-0711-47d2-89f7-ec0a85c77cd4 · inbound

SalUn: Empowering Machine Unlearning via Gradient-based Weight Saliency in Both Image Classification and Generation cites this paper.

SalUn: Empowering Machine Unlearning via Gradient-based Weight Saliency in Both Image Classification and Generation Datamodels: Predicting Predictions from Training Data

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:56:23.520684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-16T17:56:23.281678Z digest=sha256:9903854a121feeccb4871f21b7bfe2614d35d923a2dc720d389ca7c5685091e7

Observation 7bc4aac1-b63c-4eb8-b80e-aa6eb9bede2e · inbound

Merge to Mix: Mixing Datasets via Model Merging cites this paper.

Merge to Mix: Mixing Datasets via Model Merging Datamodels: Predicting Predictions from Training Data

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:09:34.826920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:09:34.826920Z digest=sha256:cc0428d5c4d02c864c76c8058bb9231b8ae294e2712305464600efd492ba53bb

Observation dbc3665d-fd3e-4707-b360-eabb2d09c7a1 · inbound

A Survey of LLM $\times$ DATA cites this paper.

A Survey of LLM $\times$ DATA Datamodels: Predicting Predictions from Training Data

Reference 184

Resolution
unresolved
no resolver link, observed 2026-08-07T14:33:13.199799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:33:13.199799Z digest=sha256:cc08df9e253fddaf90b9766c54b4175072c17e4935e02abd3d02d1b27cb5c707

Observation 8054dbf1-e20a-4a50-9820-b18105b5baf2 · inbound

Expert Survey: AI Reliability & Security Research Priorities cites this paper.

Expert Survey: AI Reliability & Security Research Priorities Datamodels: Predicting Predictions from Training Data

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:31:00.440577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:31:00.440577Z digest=sha256:4f30ee2c3fda3139d6d863f88492f2a968c40b6432eb5a19bbec6885d2cb6ec4

Observation 3ff4be62-492d-42ae-9f86-a60c48bb2f63 · inbound

Daunce: Data Attribution through Uncertainty Estimation cites this paper.

Daunce: Data Attribution through Uncertainty Estimation Datamodels: Predicting Predictions from Training Data

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:56:14.605035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:56:14.605035Z digest=sha256:d5ea0af4f77937104c93d554d505a740eb7f13345c7bf17d853361f716e4b2e2

Observation e4778dbe-e24c-4b26-8d07-f77022815ca7 · inbound

MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning cites this paper.

MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning Datamodels: Predicting Predictions from Training Data

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:19:28.727706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:19:28.727706Z digest=sha256:3659fedaab7c6fc21d26c37172944dc4126302fc2bb1deee2a933dffeb27ae5d

Observation 6415c1dd-76b9-4ad8-bea5-b82eb2c27afd · inbound

ClusterUCB: Efficient Gradient-Based Data Selection for Targeted Fine-Tuning of LLMs cites this paper.

ClusterUCB: Efficient Gradient-Based Data Selection for Targeted Fine-Tuning of LLMs Datamodels: Predicting Predictions from Training Data

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:37:07.529636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:37:07.529636Z digest=sha256:d5eae062e045f43dd6a4ede0542e5d5074d48bdd0738cb43d9c137cbb0893f0c

Observation 9bde4033-9a53-4ca5-a709-fbf69d375687 · inbound

AdaDeDup: Adaptive Hybrid Data Pruning for Efficient Large-Scale Object Detection Training cites this paper.

AdaDeDup: Adaptive Hybrid Data Pruning for Efficient Large-Scale Object Detection Training Datamodels: Predicting Predictions from Training Data

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T23:05:17.959429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:05:17.959429Z digest=sha256:1cb8527dc108f8c1d475315fa52d9d8e386707cce5f731e14cf1b6e0f8a6ce7e

Observation 14e88341-cb6e-43bd-92b4-c920351d5f2d · inbound

Scaling laws for activation steering with Llama 2 models and refusal mechanisms cites this paper.

Scaling laws for activation steering with Llama 2 models and refusal mechanisms Datamodels: Predicting Predictions from Training Data

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:05:03.982552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:05:03.982552Z digest=sha256:95289bb2a97b21c136b5c602b0bbb345ce2b0279e627368008c256fe29b49dde

Observation 81abf800-ee1b-4e0c-8b17-9521223c62bf · inbound

Better Training Data Attribution via Better Inverse Hessian-Vector Products cites this paper.

Better Training Data Attribution via Better Inverse Hessian-Vector Products Datamodels: Predicting Predictions from Training Data

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:54:42.390773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:54:42.390773Z digest=sha256:62c51e721e99436845a5ca0bcd239e14df4a8476ec1afae92d502f37095283c2

Observation 30e20088-5ed5-4df3-bb22-6921e0c7ffd8 · inbound

Understanding Data Influence with Differential Approximation cites this paper.

Understanding Data Influence with Differential Approximation Datamodels: Predicting Predictions from Training Data

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T18:28:54.951998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:28:54.951998Z digest=sha256:37e934a05a0a6fac1a09ed1478fcf63331ef33cd923ef450258439ad65600cbd

Observation b5e1f44e-e8c3-461e-899a-bf1344b81ab8 · inbound

What Is The Performance Ceiling of My Classifier? Utilizing Category-Wise Influence Functions for Pareto Frontier Analysis cites this paper.

What Is The Performance Ceiling of My Classifier? Utilizing Category-Wise Influence Functions for Pareto Frontier Analysis Datamodels: Predicting Predictions from Training Data

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T11:38:26.944685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:38:26.944685Z digest=sha256:4a827d2e25d13d16c9f74f759d999b2c69d30c5bab4e57a22dba4108115d4724

Observation 7586aea9-1aa8-49d4-89eb-a8376ae985ac · inbound

idSCD: Identifying Training Datasets through Semantic Correlation Descriptors cites this paper.

idSCD: Identifying Training Datasets through Semantic Correlation Descriptors Datamodels: Predicting Predictions from Training Data

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T08:43:15.855397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T08:35:36.928378Z digest=sha256:a03e8917e5cd09bc37b52e3aaac8da19b947853d26dbe7de237529d5aa15b347

Observation c6ee9005-6ca6-414e-9d77-ce4dc467fa48 · inbound

Quantifying the Agreement Between Data-Influence and Data-Similarity to Understand LLM Behavior cites this paper.

Quantifying the Agreement Between Data-Influence and Data-Similarity to Understand LLM Behavior Datamodels: Predicting Predictions from Training Data

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-06-26T08:49:14.925983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T08:45:34.884703Z digest=sha256:ce3bbf90176b7b02018055ca98931698abfed53966024f17c2aff90165841bbd

Observation d7a91157-3a0e-4421-bca9-15ae3499f246 · inbound

Small edits, large models: How Wikipedia advocacy shapes LLM values cites this paper.

Small edits, large models: How Wikipedia advocacy shapes LLM values Datamodels: Predicting Predictions from Training Data

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:05:36.489734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T09:03:09.977131Z digest=sha256:c86214bc1185fb7acad9c15ef461eda5f31c0bb2eaced3a7dd92856a768f5b53

Observation f2ec705d-cb1a-4c48-bb72-5fb338954e1a · inbound

Small edits, large models: How Wikipedia advocacy shapes LLM values cites this paper.

Small edits, large models: How Wikipedia advocacy shapes LLM values Datamodels: Predicting Predictions from Training Data

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-14T19:23:22.815333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T19:23:22.815333Z digest=sha256:eaa15c89a18649a47662dc53c5c1937013d1d7726f19fbb76a3efc435acacc63

Observation 245f7d70-c5be-4c5d-84c6-5b3d5853e8ba · inbound

Watermarking for Proprietary Dataset Protection cites this paper.

Watermarking for Proprietary Dataset Protection Datamodels: Predicting Predictions from Training Data

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:17:08.412539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-02T16:09:50.072430Z digest=sha256:1ef8ade675c40c1895ce7835485055a93ad435251cf679060c5b5136a2645da1

Observation 1f4e017f-4946-4d06-9464-2d5297d889e1 · inbound

An Asymptotic Analysis of the Shapley Value for Dataset Valuation cites this paper.

An Asymptotic Analysis of the Shapley Value for Dataset Valuation Datamodels: Predicting Predictions from Training Data

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T02:56:23.522332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T02:56:23.522332Z digest=sha256:3a89ebb3b6518a2261dca7e01cb0e5a2e5a4d7bcef4b52553c00902ae1d9733c

Observation 2455584f-c702-4866-9c58-3f581e4b710c · inbound

Domain-Aware Scaling Laws Uncover Data Synergy cites this paper.

Domain-Aware Scaling Laws Uncover Data Synergy Datamodels: Predicting Predictions from Training Data

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-14T07:24:27.255815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:24:27.255815Z digest=sha256:104f64477d801f93e9922dd77120878d53ea69d621eeea7e8263a5a9c53153fb