Pith. sign in

Paper Citation Record · LEDGER

In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2304.08979.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.08979 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T17:43:55.781072Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T22:15:50.062966Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 535fc342-5292-4ff5-9b4f-367d0864ba98 · inbound

"Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models cites this paper.

"Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T08:39:28.095087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-17T08:39:28.047394Z digest=sha256:217177ba051d4e827e8b8f797d5f13cb737670026625e1906a62a65eac02b7b4

Observation 868aef95-fda1-4bd6-89e9-ba937d6ff7bf · inbound

Towards Agentic Runtime Healing cites this paper.

Towards Agentic Runtime Healing In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:15:50.066128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-23T22:14:23.865122Z digest=sha256:264b2c28d321030582717b02745cbbd417792fba182a6dc9603c16ab1eea443f

Observation cda93921-c634-468f-9a54-641a68b40ad6 · inbound

Synthetic Artifact Auditing: Tracing LLM-Generated Synthetic Data Usage in Downstream Applications cites this paper.

Synthetic Artifact Auditing: Tracing LLM-Generated Synthetic Data Usage in Downstream Applications In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-09T17:43:55.781072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:43:55.781072Z digest=sha256:67d8edd7cb01b2469e94593c72406c19adc89c0367c5f5e851715f796fe38009

Observation 7d5b36e9-68f1-4530-bb59-f4f191943eed · inbound

Enhancing Health Information Retrieval with RAG by Prioritizing Topical Relevance and Factual Accuracy cites this paper.

Enhancing Health Information Retrieval with RAG by Prioritizing Topical Relevance and Factual Accuracy In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T21:59:59.195968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:59:59.195968Z digest=sha256:ebe93cb8fa1fd5c2d91f790fac163ba0a261535d8816b5e4e49b89496f4b6f67

Observation 5295d18e-ca0b-44de-904f-f43c57f26969 · inbound

InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation cites this paper.

InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:17.865016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:17.865016Z digest=sha256:f875e4ba2299c7a23f1e528d54d63a5228d9bf2c4c3c4fb81d7620850c238ce7

Observation 75a00167-735f-4d5b-84bf-38140ee83424 · inbound

From Hallucinations to Jailbreaks: Rethinking the Vulnerability of Large Foundation Models cites this paper.

From Hallucinations to Jailbreaks: Rethinking the Vulnerability of Large Foundation Models In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:34:34.621558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:34:34.621558Z digest=sha256:37765909b10cc54bebc9399b4ed766a9b46e9f9548824fd19c945cffccc5a587

Observation 0aef53cf-4632-4695-854a-f8fe3be64d87 · inbound

Bias, Accuracy, and Trust: Gender-Diverse Perspectives on Large Language Models cites this paper.

Bias, Accuracy, and Trust: Gender-Diverse Perspectives on Large Language Models In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T22:20:15.750857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:20:15.750857Z digest=sha256:0abfb2e0c10b0499a4ad1cb0c9f24dccd154eee75cfafb3cd5f2ce3c35e63757

Observation 8d116823-0417-42d2-b935-a9383d48b98b · inbound

What Shapes User Trust in ChatGPT? A Mixed-Methods Study of User Attributes, Trust Dimensions, Task Context, and Societal Perceptions among University Students cites this paper.

What Shapes User Trust in ChatGPT? A Mixed-Methods Study of User Attributes, Trust Dimensions, Task Context, and Societal Perceptions among University Students In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:50.313004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:50.313004Z digest=sha256:f46bd9e90c767b54abe026ce3bc9d2567e88e7a070dd60b7b4c64b864ea7da9a

Observation ba88cf10-a96b-4307-8f2c-e877ff3e2726 · inbound

GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs cites this paper.

GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-18T21:36:52.393660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-18T21:34:51.665401Z digest=sha256:1478a2992e6225a0706d9e5fffcc839bd9254f5632ef61d3408692cd0220f847

Observation b68836a8-d213-4490-b7d2-737b11fca782 · inbound

EPT Benchmark: Evaluation of Persian Trustworthiness in Large Language Models cites this paper.

EPT Benchmark: Evaluation of Persian Trustworthiness in Large Language Models In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T23:03:16.363007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:03:16.363007Z digest=sha256:fcdbc243d4f601030911b81e2d3a3aa4a7ce95f3e1429131d95bbd9ccd4a894c

Observation 925542b5-e7f7-4e7f-9065-b8404c6a89a2 · inbound

Red Skills or Blue Skills? A Dive Into Skills Published on ClawHub cites this paper.

Red Skills or Blue Skills? A Dive Into Skills Published on ClawHub In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:35:18.539429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-15T08:33:53.625397Z digest=sha256:f619b34a033f088c6a1c2da64c733e2540cca085523698d3ce8c5a0617dcf96e

Observation 5d49aee5-9858-43df-8f34-e3c1e4bb9678 · inbound

Discerning Authorship in Online Health Communities: Experience, Trust, and Transparency Implications for Moderating AI cites this paper.

Discerning Authorship in Online Health Communities: Experience, Trust, and Transparency Implications for Moderating AI In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:31:03.191350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T01:36:02.707490Z digest=sha256:145afed4cf173906851588221e5dde593a06eabd27e30cf44a2d23ec5207b16c

Observation 2bf5d100-886b-4f94-a008-e52f5ef534aa · inbound

Pop Quiz Attack: Black-box Membership Inference Attacks Against Large Language Models cites this paper.

Pop Quiz Attack: Black-box Membership Inference Attacks Against Large Language Models In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:26:09.159164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T09:14:12.034025Z digest=sha256:424a491422482f1e47a7500bcf90febb5564583934cccf86f0d6d419090b1fc3