Pith. sign in

Paper Citation Record · LEDGER

ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2303.15056.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.15056 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:59:22.843126Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

71
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 86697ead-1457-4f21-851c-7e73d68b855e · inbound

Visual Instruction Tuning cites this paper.

Visual Instruction Tuning ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:22:03.749556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T08:22:03.403362Z digest=sha256:8243138677adb1010ecb9491721c7292be7d183db0d987e458103eff6aeef850

Observation 39ff66de-2c9a-48ef-91c3-6229f6954662 · inbound

Enhancing Chat Language Models by Scaling High-quality Instructional Conversations cites this paper.

Enhancing Chat Language Models by Scaling High-quality Instructional Conversations ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 240

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:25:08.120802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T17:25:07.730933Z digest=sha256:3177d56c587d75598b6471838b9fa4295523d2d27f2b8dfefb3fe25923123bc0

Observation 882919d0-3361-467d-8025-cc4e9587a3b4 · inbound

Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena cites this paper.

Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T18:52:59.071717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:52:59.033645Z digest=sha256:8bd08f6df77409ec8c089968aee69b099abf046f1bb0bc8ec6f37b53e01f4c31

Observation 2b3ed6db-cbb8-4c8b-871d-9110df7a8b91 · inbound

Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning cites this paper.

Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:34:56.917191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T17:34:56.836034Z digest=sha256:84ae8daca53e36b836f287ae84d6d7bd2f48161ba32b07bd65b5a86162e280a3

Observation 14cd40a8-3af4-452f-8b6b-c0d0e0a491e1 · inbound

RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback cites this paper.

RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 90

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:32:28.019478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T21:32:27.806494Z digest=sha256:a90cfe22c93bf0f011670e47c0efd36c2f38377feb0a0cde89f7f9523320beb0

Observation df608a20-961d-4a1a-a590-0248b4ba1c00 · inbound

Chain-of-Verification Reduces Hallucination in Large Language Models cites this paper.

Chain-of-Verification Reduces Hallucination in Large Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 140

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:06:50.336442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T01:06:49.811982Z digest=sha256:1432f88548596c25a01b363ab090e7c0afc42fe6b302257dac383d196f4da2c0

Observation 6692d942-ac0f-47e3-bdc8-a178f2e22c9d · inbound

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models cites this paper.

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 287

Resolution
verified exact
arxiv_id, observed 2026-05-18T06:38:37.160208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T06:38:36.517935Z digest=sha256:8ab00e2cdf19913578b6eec6ea9dccf6582e50d6588e9b9efa54fe79fc6eeca2

Observation e74d8083-0eb7-4527-8b75-5757ca1d7b26 · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-17T12:04:10.577423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:a140525eb233fca6e3dae9f1963871c98a843ab102cb635581f8d35e8ea07358

Observation c2d74b0e-2cfc-490e-b576-61cb250bd370 · inbound

Simple Prompt Injection Attacks Can Leak Personal Data Observed by LLM Agents During Task Execution cites this paper.

Simple Prompt Injection Attacks Can Leak Personal Data Observed by LLM Agents During Task Execution ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:22.843126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:22.843126Z digest=sha256:a2571abf9e3df1c1a427b37c3e53142f232894d94225b871408503edd1658c57

Observation 5e8af0c7-399c-4a34-a539-655d3c378448 · inbound

Prompt Candidates, then Distill: A Teacher-Student Framework for LLM-driven Data Annotation cites this paper.

Prompt Candidates, then Distill: A Teacher-Student Framework for LLM-driven Data Annotation ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T11:01:31.478037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:01:31.478037Z digest=sha256:99c42603e289e51e4ff79f32bfbb9fc5775ff63e777e2ef37d50bb8887c14fc2

Observation 05ffb229-30b1-4835-97db-4bc9ddd22642 · inbound

Towards Efficient and Effective Alignment of Large Language Models cites this paper.

Towards Efficient and Effective Alignment of Large Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T04:55:39.571857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:55:39.571857Z digest=sha256:39bb03b8ec5757cde9ec91a58c55e19a04fc971705e82ea0b292381060b40169

Observation 4de7c867-4b92-4e56-8f9f-f55197ffec7d · inbound

Using AI to replicate human experimental results: a motion study cites this paper.

Using AI to replicate human experimental results: a motion study ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T17:37:30.227785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:37:30.227785Z digest=sha256:9d6e7afc9edf231a716cfa5deaa833825e3918e167084b5ec2b14bc5bb8eaf11

Observation c2b645b4-8125-4f13-aa1f-d0ae38348701 · inbound

Just Put a Human in the Loop? Investigating LLM-Assisted Annotation for Subjective Tasks cites this paper.

Just Put a Human in the Loop? Investigating LLM-Assisted Annotation for Subjective Tasks ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:04.188426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:04.188426Z digest=sha256:8628600ac8b670800d30f2167e8f5a1820d64776ba57b1afad4edba24e0ea7e8

Observation 6cf5b94e-3daf-44dc-8c0d-f10f6e1250aa · inbound

Guidelines for Empirical Studies in Software Engineering involving Large Language Models cites this paper.

Guidelines for Empirical Studies in Software Engineering involving Large Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:02:52.213502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T22:02:36.307598Z digest=sha256:33017fba75601a229e2617b9ceebeb8263ebb51ab3f543267ac4deaa1e54b63d

Observation 5887031b-d97b-4b9b-b220-606a97bca22b · inbound

Guidelines for Empirical Studies in Software Engineering involving Large Language Models cites this paper.

Guidelines for Empirical Studies in Software Engineering involving Large Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:20:31.709513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T08:18:18.448122Z digest=sha256:cd71b62ebe69ff0c6ed9b31c043ad7b5b45807a86227e39d6967a1c779dcab57

Observation 7167ba93-7cdb-48c7-be19-2973bd35ae55 · inbound

Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection cites this paper.

Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:40:27.025055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T13:38:02.450128Z digest=sha256:ed54b487ad8dce7b5283cb86ff89d3979eb60f3679e57812e9bfa51b897791ae

Observation cfc0df83-aab3-43b8-861d-d3db85060f38 · inbound

Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection cites this paper.

Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T20:35:52.598345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:35:52.598345Z digest=sha256:37755551d13b0b6de4fb7a978e394aae0290458907564f694c5525e801a04fb2

Observation f26bc985-4ba5-4c0a-8145-b2e5de255ef6 · inbound

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages cites this paper.

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 150

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:38:21.738599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T14:33:36.100966Z digest=sha256:64dc65972c93f5ae2d07b03a7d20261a3613354fcace8ee3f7102ae2cffba3bc

Observation 2a0ad235-2a2f-4444-bec6-1eaa18c479a0 · inbound

Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization cites this paper.

Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T13:13:27.570712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T13:05:15.701729Z digest=sha256:2af4faba5251d601179ce5a4dbde22a6ecb0aa9591b62ee7beb246353a6a7aab

Observation 4f8d98ad-81b0-416e-be19-0a571b5ebf90 · inbound

Structure Before Collapse: Transient semantic geometry in next-token prediction cites this paper.

Structure Before Collapse: Transient semantic geometry in next-token prediction ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:29:51.131669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T05:14:07.208255Z digest=sha256:601794f93192b6065de1d14c2426be5c40b9b49ac8bd0a10acb7a0523c599cec

Observation b2c27081-31b0-4912-b352-d4f76c96a608 · inbound

CORTEX: High-Quality Cross-Domain Organization of Web-Scale Corpora through Ontological Corpus Graph cites this paper.

CORTEX: High-Quality Cross-Domain Organization of Web-Scale Corpora through Ontological Corpus Graph ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:14:18.532747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-30T06:04:40.684934Z digest=sha256:71134e74d92f293d9f7cd19dc4f68b40fae3342f6346dd89d5ed85ad11663f95

Observation 1c8a3931-ef25-47f7-8f0d-3cb8305b55e4 · inbound

SERUM: State Extraction and Refinement for User Modeling cites this paper.

SERUM: State Extraction and Refinement for User Modeling ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T12:13:37.936172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:13:37.936172Z digest=sha256:7cb99ac28acb5dd95116f650f04c35525b674219824bc39a0211dcec27a08d28