Pith. sign in

Paper Citation Record · LEDGER

In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2304.08979.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.08979 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T17:43:55.781072Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T22:15:50.062966Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 535fc342-5292-4ff5-9b4f-367d0864ba98 · inbound

"Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models cites this paper.

"Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T08:39:28.095087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T08:39:28.047394Z digest=sha256:cd6c4e2dd66c3655f4658dafc6d5870096a8c862013a9251cca99d745a2457d9

Observation 868aef95-fda1-4bd6-89e9-ba937d6ff7bf · inbound

Towards Agentic Runtime Healing cites this paper.

Towards Agentic Runtime Healing In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:15:50.066128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T22:14:23.865122Z digest=sha256:ffaaab013b0cbfdf7b826862a3001f734e1a4721c93ffbec43ea0f876f38af23

Observation cda93921-c634-468f-9a54-641a68b40ad6 · inbound

Synthetic Artifact Auditing: Tracing LLM-Generated Synthetic Data Usage in Downstream Applications cites this paper.

Synthetic Artifact Auditing: Tracing LLM-Generated Synthetic Data Usage in Downstream Applications In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-09T17:43:55.781072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:43:55.781072Z digest=sha256:4af2d4bc7fbb38ad98e4127065a49a5ed6734996ee4e2a669560698c34e3cae3

Observation 7d5b36e9-68f1-4530-bb59-f4f191943eed · inbound

Enhancing Health Information Retrieval with RAG by Prioritizing Topical Relevance and Factual Accuracy cites this paper.

Enhancing Health Information Retrieval with RAG by Prioritizing Topical Relevance and Factual Accuracy In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T21:59:59.195968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:59:59.195968Z digest=sha256:000b7a26378e536a876c072a3919e265b99c8cfd4881902701b3447bfb3903f1

Observation 5295d18e-ca0b-44de-904f-f43c57f26969 · inbound

InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation cites this paper.

InfoDeepSeek: Benchmarking Agentic Information Seeking for Retrieval-Augmented Generation In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:17.865016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:17.865016Z digest=sha256:7a60daee04eeec7d4a3aa7d9c8c2ce2db8243cbfa2b1f95fbf6ff531e48cea75

Observation 75a00167-735f-4d5b-84bf-38140ee83424 · inbound

From Hallucinations to Jailbreaks: Rethinking the Vulnerability of Large Foundation Models cites this paper.

From Hallucinations to Jailbreaks: Rethinking the Vulnerability of Large Foundation Models In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:34:34.621558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:34:34.621558Z digest=sha256:5de2bb799be180c94eafb81013010f00b634873b30d7305589211aa81b328cf0

Observation 0aef53cf-4632-4695-854a-f8fe3be64d87 · inbound

Bias, Accuracy, and Trust: Gender-Diverse Perspectives on Large Language Models cites this paper.

Bias, Accuracy, and Trust: Gender-Diverse Perspectives on Large Language Models In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T22:20:15.750857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:20:15.750857Z digest=sha256:553980ee6064509817d0930c750f03afb54506be47671b7dd7794ba9a32ce762

Observation 8d116823-0417-42d2-b935-a9383d48b98b · inbound

What Shapes User Trust in ChatGPT? A Mixed-Methods Study of User Attributes, Trust Dimensions, Task Context, and Societal Perceptions among University Students cites this paper.

What Shapes User Trust in ChatGPT? A Mixed-Methods Study of User Attributes, Trust Dimensions, Task Context, and Societal Perceptions among University Students In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:50.313004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:50.313004Z digest=sha256:f81f6e8322cd0c8086a2c3e278047c4020afef89685a15ce97077d0704a7739c

Observation ba88cf10-a96b-4307-8f2c-e877ff3e2726 · inbound

GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs cites this paper.

GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-18T21:36:52.393660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T21:34:51.665401Z digest=sha256:0f83be1ad6f02dc70895c59c3df492a8fbe6c44834f179be2d8dc25ba91c3682

Observation b68836a8-d213-4490-b7d2-737b11fca782 · inbound

EPT Benchmark: Evaluation of Persian Trustworthiness in Large Language Models cites this paper.

EPT Benchmark: Evaluation of Persian Trustworthiness in Large Language Models In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T23:03:16.363007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:03:16.363007Z digest=sha256:e1ca5604a17a7490c93b90f6a2c03a7f3c4b0f1f58d5de90ccc4f54916de8ff2

Observation 925542b5-e7f7-4e7f-9065-b8404c6a89a2 · inbound

Red Skills or Blue Skills? A Dive Into Skills Published on ClawHub cites this paper.

Red Skills or Blue Skills? A Dive Into Skills Published on ClawHub In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:35:18.539429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T08:33:53.625397Z digest=sha256:014b70719c7da488db5a5b203f696a0a99317a47f398b769719c5feef05110a5

Observation 5d49aee5-9858-43df-8f34-e3c1e4bb9678 · inbound

Discerning Authorship in Online Health Communities: Experience, Trust, and Transparency Implications for Moderating AI cites this paper.

Discerning Authorship in Online Health Communities: Experience, Trust, and Transparency Implications for Moderating AI In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:31:03.191350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T01:36:02.707490Z digest=sha256:4c761364913d01d57109e4daf66ca84e0a2b80c18535920a1958ccc066035c0d

Observation 2bf5d100-886b-4f94-a008-e52f5ef534aa · inbound

Pop Quiz Attack: Black-box Membership Inference Attacks Against Large Language Models cites this paper.

Pop Quiz Attack: Black-box Membership Inference Attacks Against Large Language Models In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:26:09.159164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T09:14:12.034025Z digest=sha256:044bdb75c3c83291bddb11038b408717f35c04acf30724e68fc73eacd21863cb