Pith. sign in

Paper Citation Record · LEDGER

ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2303.15056.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.15056 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:59:22.843126Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

71
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 86697ead-1457-4f21-851c-7e73d68b855e · inbound

Visual Instruction Tuning cites this paper.

Visual Instruction Tuning ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:22:03.749556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T08:22:03.403362Z digest=sha256:57fe9c576983e42c156d8c0d0e61b4d135737421eb97dab4d582df7db78f2762

Observation 39ff66de-2c9a-48ef-91c3-6229f6954662 · inbound

Enhancing Chat Language Models by Scaling High-quality Instructional Conversations cites this paper.

Enhancing Chat Language Models by Scaling High-quality Instructional Conversations ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 240

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:25:08.120802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-15T17:25:07.730933Z digest=sha256:e21063b7ade5ce520b89121957a478ea39154b827ea199a82ab5ffde914c93c9

Observation 882919d0-3361-467d-8025-cc4e9587a3b4 · inbound

Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena cites this paper.

Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T18:52:59.071717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T18:52:59.033645Z digest=sha256:2c43e73b142bfb83c6919d65b467dce11ed4a528b3c6149ada556176680bb08a

Observation 2b3ed6db-cbb8-4c8b-871d-9110df7a8b91 · inbound

Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning cites this paper.

Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:34:56.917191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T17:34:56.836034Z digest=sha256:73b3a015a50ef9e3cac67cbc383e74830dc1770eaf40b5bd70e013915b79e70e

Observation 14cd40a8-3af4-452f-8b6b-c0d0e0a491e1 · inbound

RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback cites this paper.

RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 90

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:32:28.019478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-15T21:32:27.806494Z digest=sha256:5b25f4ca1dd11decc99878c2fa7a3b91a9bdbfbb31a8653f319f7379cce12550

Observation df608a20-961d-4a1a-a590-0248b4ba1c00 · inbound

Chain-of-Verification Reduces Hallucination in Large Language Models cites this paper.

Chain-of-Verification Reduces Hallucination in Large Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 140

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:06:50.336442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T01:06:49.811982Z digest=sha256:75437af5c70c6a855a210dd68c8cd9e9fe6276768e0d8b4668246d5f746cdc7c

Observation 6692d942-ac0f-47e3-bdc8-a178f2e22c9d · inbound

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models cites this paper.

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 287

Resolution
verified exact
arxiv_id, observed 2026-05-18T06:38:37.160208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T06:38:36.517935Z digest=sha256:9e2788d991f950b0ae8237d5888383cc937917dd63d86bbb3bc37d258c86c298

Observation e74d8083-0eb7-4527-8b75-5757ca1d7b26 · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-17T12:04:10.577423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:c344c115a67b85cd5ff3ca6d4c90020d6a4204980abea2a7c4f1cc89bec160ac

Observation c2d74b0e-2cfc-490e-b576-61cb250bd370 · inbound

Simple Prompt Injection Attacks Can Leak Personal Data Observed by LLM Agents During Task Execution cites this paper.

Simple Prompt Injection Attacks Can Leak Personal Data Observed by LLM Agents During Task Execution ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:22.843126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:22.843126Z digest=sha256:a2571abf9e3df1c1a427b37c3e53142f232894d94225b871408503edd1658c57

Observation 5e8af0c7-399c-4a34-a539-655d3c378448 · inbound

Prompt Candidates, then Distill: A Teacher-Student Framework for LLM-driven Data Annotation cites this paper.

Prompt Candidates, then Distill: A Teacher-Student Framework for LLM-driven Data Annotation ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T11:01:31.478037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:01:31.478037Z digest=sha256:99c42603e289e51e4ff79f32bfbb9fc5775ff63e777e2ef37d50bb8887c14fc2

Observation 05ffb229-30b1-4835-97db-4bc9ddd22642 · inbound

Towards Efficient and Effective Alignment of Large Language Models cites this paper.

Towards Efficient and Effective Alignment of Large Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T04:55:39.571857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:55:39.571857Z digest=sha256:39bb03b8ec5757cde9ec91a58c55e19a04fc971705e82ea0b292381060b40169

Observation 4de7c867-4b92-4e56-8f9f-f55197ffec7d · inbound

Using AI to replicate human experimental results: a motion study cites this paper.

Using AI to replicate human experimental results: a motion study ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T17:37:30.227785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:37:30.227785Z digest=sha256:9d6e7afc9edf231a716cfa5deaa833825e3918e167084b5ec2b14bc5bb8eaf11

Observation c2b645b4-8125-4f13-aa1f-d0ae38348701 · inbound

Just Put a Human in the Loop? Investigating LLM-Assisted Annotation for Subjective Tasks cites this paper.

Just Put a Human in the Loop? Investigating LLM-Assisted Annotation for Subjective Tasks ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T15:27:04.188426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:27:04.188426Z digest=sha256:8628600ac8b670800d30f2167e8f5a1820d64776ba57b1afad4edba24e0ea7e8

Observation 6cf5b94e-3daf-44dc-8c0d-f10f6e1250aa · inbound

Guidelines for Empirical Studies in Software Engineering involving Large Language Models cites this paper.

Guidelines for Empirical Studies in Software Engineering involving Large Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:02:52.213502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T22:02:36.307598Z digest=sha256:7f54cd8e588f0688d4753b16bb7cbe289f0af3022ce59a8c2732c337046e29f1

Observation 5887031b-d97b-4b9b-b220-606a97bca22b · inbound

Guidelines for Empirical Studies in Software Engineering involving Large Language Models cites this paper.

Guidelines for Empirical Studies in Software Engineering involving Large Language Models ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:20:31.709513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T08:18:18.448122Z digest=sha256:39d97e474fe8965878b5f893616d00dd62ead725c94a2195c49a643d71295a4f

Observation 7167ba93-7cdb-48c7-be19-2973bd35ae55 · inbound

Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection cites this paper.

Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T13:40:27.025055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T13:38:02.450128Z digest=sha256:13b544c3a6954bc289569b8b5fe2b16e506ee78b91975c03dcc4fed6b8eb7599

Observation cfc0df83-aab3-43b8-861d-d3db85060f38 · inbound

Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection cites this paper.

Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T20:35:52.598345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:35:52.598345Z digest=sha256:37755551d13b0b6de4fb7a978e394aae0290458907564f694c5525e801a04fb2

Observation f26bc985-4ba5-4c0a-8145-b2e5de255ef6 · inbound

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages cites this paper.

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 150

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:38:21.738599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-20T14:33:36.100966Z digest=sha256:567e60d8ad9fbe335048ef0546e2a7afd439e6e5003ad59473b098dd9ecad5a4

Observation 2a0ad235-2a2f-4444-bec6-1eaa18c479a0 · inbound

Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization cites this paper.

Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T13:13:27.570712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T13:05:15.701729Z digest=sha256:409c4b5ea5e7a0b7951633beb7ce4a8fe6396b65ad9dd8098fec0d8abf4c2d4c

Observation 4f8d98ad-81b0-416e-be19-0a571b5ebf90 · inbound

Structure Before Collapse: Transient semantic geometry in next-token prediction cites this paper.

Structure Before Collapse: Transient semantic geometry in next-token prediction ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:29:51.131669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T05:14:07.208255Z digest=sha256:4ed7c34141339eeec81d52dac5212cb6ee7175e96196d21918b495f15b6457f5

Observation b2c27081-31b0-4912-b352-d4f76c96a608 · inbound

CORTEX: High-Quality Cross-Domain Organization of Web-Scale Corpora through Ontological Corpus Graph cites this paper.

CORTEX: High-Quality Cross-Domain Organization of Web-Scale Corpora through Ontological Corpus Graph ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:14:18.532747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-30T06:04:40.684934Z digest=sha256:c134ccf2bdc7d47670a0b172ec94ffcc6656125ed238d139bc44df4ded56d3ea

Observation 1c8a3931-ef25-47f7-8f0d-3cb8305b55e4 · inbound

SERUM: State Extraction and Refinement for User Modeling cites this paper.

SERUM: State Extraction and Refinement for User Modeling ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T12:13:37.936172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:13:37.936172Z digest=sha256:7cb99ac28acb5dd95116f650f04c35525b674219824bc39a0211dcec27a08d28