Pith. sign in

Paper Citation Record · LEDGER

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis

As of 18 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 2 inbound Pith citation observations for arXiv:2505.03019.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.03019 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:07:34.103309Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:46:43.991193Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T21:11:50.756949Z

Reference resolution

20 of 20 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 24695354-4742-42e9-b906-b71d8acad233 · outbound

This paper cites https://www.europarl.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis https://www.europarl

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.029898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.029898Z digest=sha256:013b1afe99207182f8729ff3f61d13b81b8d0442b77b2b072b419b567017a2a9

Observation 74e7b0e7-b359-4b1f-939b-e5db3fc9da6b · outbound

This paper cites Quantifying Memorization Across Neural Language Models.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis Quantifying Memorization Across Neural Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.038036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.038036Z digest=sha256:56e997ab0c560f3e677a9f00112b65c51bf34be0b00b7e5c7f4e2e1a526797d9

Observation c02fa179-cee8-4225-809f-bd6510e53573 · outbound

This paper cites Generalization or Memorization: Data Contamination and Trustworthy Evaluation for Large Language Models.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis Generalization or Memorization: Data Contamination and Trustworthy Evaluation for Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.046125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.046125Z digest=sha256:2ab7baae729555a60bbd59f458563c8d99a8b399e0d9a64fd5ff7bcd7d79d951

Observation 2d90c191-cc0c-4928-98fd-5be4986b5eec · outbound

This paper cites Do Membership Inference Attacks Work on Large Language Models?.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis Do Membership Inference Attacks Work on Large Language Models?

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.049906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.049906Z digest=sha256:5afc2002564be41dbf817b7affb4be6e7f0f6004473cd717268fb898056253e9

Observation 491d7ee5-ea69-436e-9c82-ae585f9333f4 · outbound

This paper cites Time Travel in LLMs: Tracing Data Contamination in Large Language Models.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis Time Travel in LLMs: Tracing Data Contamination in Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.058192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.058192Z digest=sha256:b7b12e9e834edc839d631b4babc014eb396acc11c5947a9e29d9bc80fc5b1d8b

Observation 25f11f27-2646-4ece-9566-77cc24e1ba57 · outbound

This paper cites SoK: Memorization in General-Purpose Large Language Models.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis SoK: Memorization in General-Purpose Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.062031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.062031Z digest=sha256:7411bc7ca575729f4cac7682faf938f6203dbaccb905ca37ca4d82b1f2ec35b5

Observation 085a9689-365d-48d6-8de6-3c3d82ded7e8 · outbound

This paper cites org/abstract/document/10179300/.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis org/abstract/document/10179300/

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.069793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.069793Z digest=sha256:f8fc12c6ec7e2d314398a349286fcb3e099ef71fca83d8573ceddad2b1c8fa54

Observation f0025452-bc9f-4144-9d08-b75b6aa59c95 · outbound

This paper cites On Leakage of Code Generation Evaluation Datasets.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis On Leakage of Code Generation Evaluation Datasets

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.073640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.073640Z digest=sha256:54a592b4c798b8184ba6be99ceb81463674525dd3197a8639bc74238756eb4b3

Observation 444a8571-d1cf-4bfb-be2a-b42fa80f90ec · outbound

This paper cites Penedo, G., Malartic, Q., Hesslow, D., Cojocaru, R., Cap- pelli, A., Alobeidli, H., Pannier, B., Almazrouei, E., and Launay, J.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis Penedo, G., Malartic, Q., Hesslow, D., Cojocaru, R., Cap- pelli, A., Alobeidli, H., Pannier, B., Almazrouei, E., and Launay, J

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:07:34.439529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T00:07:34.077528Z digest=sha256:7047d56709ccab854f87fed1bb3882de9b84553746acc46d65ee9c6b67e3f406

Observation 98f64cc9-4df8-4990-bbb0-b131bb0bb45f · outbound

This paper cites The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.081143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.081143Z digest=sha256:6cad68705606edaab4715ae2d088847750a0f6a8cd2a0642c311df9571518e7f

Observation d180c1d1-eba5-4cda-bef7-eec595e4d753 · outbound

This paper cites Rethinking LLM Memorization through the Lens of Adversarial Compression.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis Rethinking LLM Memorization through the Lens of Adversarial Compression

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.084901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.084901Z digest=sha256:9552f84023bb7ddb0173e94d8201047d9d1c35133c02ea5bf9d59e0d762950ef

Observation e96eacc9-41e3-4134-bb49-27e919b260a8 · outbound

This paper cites Understanding Memorisation in LLMs: Dynamics, Influencing Factors, and Implications.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis Understanding Memorisation in LLMs: Dynamics, Influencing Factors, and Implications

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.088695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.088695Z digest=sha256:2f5c3ca1af2fc505c87e679c090f40cbfa832a3d0fc4d916ce5ec98efa78dd3f

Observation ea11e795-eac1-4643-bc1b-a7a1fb70affd · outbound

This paper cites On Protecting the Data Privacy of Large Language Models (LLMs): A Survey.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis On Protecting the Data Privacy of Large Language Models (LLMs): A Survey

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.092304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.092304Z digest=sha256:0ce63320ae5163ce5dd1c702a47a46af75e2319288fc30612d1bd9be8ec6ac64

Observation 921f573c-2f2e-473b-be9e-e9241d3e29ca · outbound

This paper cites Publisher: Elsevier.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis Publisher: Elsevier

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:07:34.425889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T00:07:34.095960Z digest=sha256:d143fecc31fbbfb223b8f29ac410affde12ad68a16bf6d4db89b913710c9bf6f

Observation 036a795e-a5ba-413c-a70f-17d1db8a4ebe · outbound

This paper cites Don't Make Your LLM an Evaluation Benchmark Cheater.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis Don't Make Your LLM an Evaluation Benchmark Cheater

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.099417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.099417Z digest=sha256:f5624979942acc6143709e8c18cc55450b0aced4c053a59ee7b1b78172cf0903

Observation 9a0de113-e215-40cb-aa66-11fe678d4a03 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.103309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.103309Z digest=sha256:c2b8cd3ab139565d6c7554da61a9633840db5eb010752ec19dc22741edcd1e47

Observation ada311f7-a6cb-49ca-8d9e-602c03b01b95 · outbound

This paper cites URL https://aclanthology.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis URL https://aclanthology

Reference 2004

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T00:07:34.450700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T00:07:34.065847Z digest=sha256:ea4051e905d83008cee7b834e6fb71249e53adcd1d69e69d0010e934a587bc09

Observation be430733-5d0a-497a-ae94-26ba07ce04e8 · outbound

This paper cites The Pile: An 800GB Dataset of Diverse Text for Language Modeling.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis The Pile: An 800GB Dataset of Diverse Text for Language Modeling

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.054210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.054210Z digest=sha256:86ea302b1b6ab875068538a95733a9b3ebb45a2385c4099d34c3f55ece07d14d

Observation 3e603aed-f286-40e8-aa0c-d22f691fec76 · outbound

This paper cites Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-16T00:07:34.034035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:07:34.034035Z digest=sha256:c86b8a9fc9f94fdec45a6a9cfb570cbeb4a17d2ae67f8ae582dc7c76f70d65c8

Observation 534eddad-b28b-491d-ba82-b0291ed9c0b3 · outbound

This paper cites Generalisation First, Memorisation Second? Memorisation Localisation for Natural Language Classification Tasks.

Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis Generalisation First, Memorisation Second? Memorisation Localisation for Natural Language Classification Tasks

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-16T00:07:34.341823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T00:07:34.041971Z digest=sha256:ffa026c5d00ad31c8e63213dd3d62276cc1de10e760a44b47a75386dd66d8406

Pith citing papers

Observation 92948611-ba40-4484-ba14-c17e5cccf686 · inbound

Test of Time: Rethinking Temporal Signal of Benchmark Contamination cites this paper.

Test of Time: Rethinking Temporal Signal of Benchmark Contamination Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-18T21:11:50.760908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-18T21:10:22.237418Z digest=sha256:244404f5ee4b26249e122ce2d4c8789bf13e48177bd8a2a306e3f05570fd72d3

Observation 62030fa9-e39b-4e18-a8a5-d71c4dff1972 · inbound

Memorization Diagnostics for Code LLMs Should be Scale-Aware cites this paper.

Memorization Diagnostics for Code LLMs Should be Scale-Aware Memorization or Interpolation ? Detecting LLM Memorization through Input Perturbation Analysis

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:43.991193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:43.991193Z digest=sha256:426f2a469785718813f36ecd69e4f9277dc5929446dcde683afc5e67c81cb7a1