Pith. sign in

Paper Citation Record · LEDGER

ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2403.16887.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.16887 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:36:50.435273Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T19:46:10.780201Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 41dce68e-add3-4b04-9581-2ef7c7ba5725 · inbound

Why Does ChatGPT "Delve" So Much? Exploring the Sources of Lexical Overrepresentation in Large Language Models cites this paper.

Why Does ChatGPT "Delve" So Much? Exploring the Sources of Lexical Overrepresentation in Large Language Models ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T15:02:39.519108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:02:39.519108Z digest=sha256:eee3e7e206f327efbfceb2281a501e3c71afa22628b2860b7de888f423fd714c

Observation b605a175-08ed-4216-a09f-e206eef1bf84 · inbound

NLLG Quarterly arXiv Report 09/24: What are the most influential current AI Papers? cites this paper.

NLLG Quarterly arXiv Report 09/24: What are the most influential current AI Papers? ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T00:03:53.731093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:03:53.731093Z digest=sha256:2c864532b44193c0f46219080467cd7afabdb6e5ca7cacae5563ef4dd870dba9

Observation 6da427d6-a6f7-4f1e-b8c6-a38ea03f4c89 · inbound

Human-LLM Coevolution: Evidence from Academic Writing cites this paper.

Human-LLM Coevolution: Evidence from Academic Writing ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T20:54:17.430756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T20:54:17.430756Z digest=sha256:b7b0d39c690e7f7f2c9b8416dd30ed3d1325670edf5a798d031618ac02bfadad

Observation af174f94-43fd-4e9e-b5ef-7f03a350a246 · inbound

Large Means Left: Political Bias in Large Language Models Increases with Their Number of Parameters cites this paper.

Large Means Left: Political Bias in Large Language Models Increases with Their Number of Parameters ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T23:36:50.435273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:36:50.435273Z digest=sha256:42453c75fb61e36b454b26cd8209599b44286fea05d8a443b98a3eeda6fae494

Observation dd234b9b-5631-4afb-9a2f-fb0d9161034b · inbound

GPT Editors, Not Authors: The Stylistic Footprint of LLMs in Academic Preprints cites this paper.

GPT Editors, Not Authors: The Stylistic Footprint of LLMs in Academic Preprints ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:16.366835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:16.366835Z digest=sha256:9aa25fd37ed2c528e293bb21f1973f81a0c6f2aa29cc4c1bdfec9c804182099b

Observation aeb8e5c2-c365-4143-89fb-ce6e02fa471b · inbound

Exploring the Structure of AI-Induced Language Change in Scientific English cites this paper.

Exploring the Structure of AI-Induced Language Change in Scientific English ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T22:24:08.054203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:24:08.054203Z digest=sha256:a49605d48b394a95c626b8199ddc81b63d21828645fde29d1adcfdb03bbfa6b7

Observation f27c1ca3-6b13-4f45-be88-4c749110f3de · inbound

Model Misalignment and Language Change: Traces of AI-Associated Language in Unscripted Spoken English cites this paper.

Model Misalignment and Language Change: Traces of AI-Associated Language in Unscripted Spoken English ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T10:23:57.773931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:23:57.773931Z digest=sha256:e376b0b64f7cd8856f50c286cf70383ed94f98e75a56af9cdefd6303e36c6465

Observation 1d744fbb-99c1-4e68-9549-abd0021ed408 · inbound

Word Overuse and Alignment in Large Language Models: The Influence of Learning from Human Feedback cites this paper.

Word Overuse and Alignment in Large Language Models: The Influence of Learning from Human Feedback ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T05:23:13.243058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:23:13.243058Z digest=sha256:3840afbf1a69904e374ef23a152704159491790a6e1d76f41670492e718dc1d0

Observation 0c0a73bc-2940-4a3e-8f85-02594ae79ac0 · inbound

What Are LLMs Doing to Scientific Communication? Measuring Changes in Writing Practices and Reading Experience cites this paper.

What Are LLMs Doing to Scientific Communication? Measuring Changes in Writing Practices and Reading Experience ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:08:04.955010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T06:06:12.273910Z digest=sha256:e387d2d22499060280ba1ff8d116e4a066ce937179b684a3cec57a1831276915

Observation 38eb999d-21f8-400f-ae23-ac760dd8be5e · inbound

PRAIB: Peer Review AI Benchmark of Behaviour of LLM-Assisted Reviewing cites this paper.

PRAIB: Peer Review AI Benchmark of Behaviour of LLM-Assisted Reviewing ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:23:13.104534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T07:16:14.345586Z digest=sha256:ffcd6f2b899b9ea56efde5812399a8c378378974330f0a4839d2936f377caa5a

Observation f5c37277-c932-4768-8c8e-f1c8c50c398a · inbound

Isolating LLM Lexical Bias: A Curation-Free Triangulated Metric for Preference-Stage Learning cites this paper.

Isolating LLM Lexical Bias: A Curation-Free Triangulated Metric for Preference-Stage Learning ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T19:46:10.781710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-28T22:04:42.187726Z digest=sha256:01f2802d5a3dbfe917d49808ffc7f4476684e6a3a3f4d55c7084582e4b7f3ff4

Observation 15f75325-9194-44f9-afd6-fcf187fecc54 · inbound

Most biomedical publications show signs of LLM-assisted writing cites this paper.

Most biomedical publications show signs of LLM-assisted writing ChatGPT "contamination": estimating the prevalence of LLMs in the scholarly literature

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T18:54:18.758716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:54:18.758716Z digest=sha256:0136cc5de34e152bff5053067bd523fa521d4dc3aba8031668ee24cb00180285