Pith. sign in

Paper Citation Record · LEDGER

Human-like Summarization Evaluation with ChatGPT

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2304.02554.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.02554 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:19:34.827221Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

38
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 67ef8c81-3b94-4d19-84a7-3575f5e7b04d · inbound

ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate cites this paper.

ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate Human-like Summarization Evaluation with ChatGPT

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-13T13:03:18.810777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T13:03:18.765496Z digest=sha256:06a8fdb20a20138a6711373593a0e89b40805122b97c88bd1e47c8d10072aa4c

Observation 3ef92e9e-ed7a-4170-8e6b-ff61fae66aea · inbound

A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions cites this paper.

A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions Human-like Summarization Evaluation with ChatGPT

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:46:27.555802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T02:46:26.957539Z digest=sha256:2916b452b26941177fee251426ec0567fc93172b3a9afbe9809e1d90ee6ecf30

Observation 902dbcb4-60da-4398-829b-1baadea826d7 · inbound

AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions cites this paper.

AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions Human-like Summarization Evaluation with ChatGPT

Reference 228

Resolution
verified exact
arxiv_id, observed 2026-05-23T21:55:49.997989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T21:54:26.670284Z digest=sha256:57c3200850292c67f2fc9cca95227e65bb4323370f2a86e71ee0963181706ded

Observation 06000e1c-39a1-4252-8da1-202daa152937 · inbound

A Survey on LLM-as-a-Judge cites this paper.

A Survey on LLM-as-a-Judge Human-like Summarization Evaluation with ChatGPT

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:35:44.287787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T17:33:13.394338Z digest=sha256:6aaff130cc5eb04bee0176a493eaef0d37d6c506722e0ee6f5436de29359f9a3

Observation 7ed344d8-cc99-4521-8a1f-236530c45dda · inbound

Efficient Online RFT with Plug-and-Play LLM Judges: Unlocking State-of-the-Art Performance cites this paper.

Efficient Online RFT with Plug-and-Play LLM Judges: Unlocking State-of-the-Art Performance Human-like Summarization Evaluation with ChatGPT

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T10:19:34.827221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:19:34.827221Z digest=sha256:6a90482731ef6f149296274f581b74ceb4571fee3782e2048d9e2f15488e9752

Observation 6d4cbf06-a589-4f36-8b55-fe83b1168291 · inbound

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation cites this paper.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Human-like Summarization Evaluation with ChatGPT

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.575596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.575596Z digest=sha256:c1e2ba685e908d7593b04bd5ab42aa417594e401e67161e2237adbbad299d8f4

Observation 741f0943-23d7-4e97-a828-35f987c12e2a · inbound

Reliable Annotations with Less Effort: Evaluating LLM-Human Collaboration in Search Clarifications cites this paper.

Reliable Annotations with Less Effort: Evaluating LLM-Human Collaboration in Search Clarifications Human-like Summarization Evaluation with ChatGPT

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:31.638854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:31.638854Z digest=sha256:7be0f7b7d6873a12615bb9f543c709421d7d6f259d0ea67f8eec22050f76c648

Observation 94ea1d46-cb64-4f75-8c1d-946df320e2c7 · inbound

Byzantine-Robust Decentralized Coordination of LLM Agents cites this paper.

Byzantine-Robust Decentralized Coordination of LLM Agents Human-like Summarization Evaluation with ChatGPT

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:01.188665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:01.188665Z digest=sha256:0a9757fd446e3c45befb7eddc12d6400089be07b1f32551b9ac1c7e9f372885d

Observation 4fe04d12-60a9-4794-9cab-1de6c3c9ada0 · inbound

Play Favorites: A Statistical Method to Measure Self-Bias in LLM-as-a-Judge cites this paper.

Play Favorites: A Statistical Method to Measure Self-Bias in LLM-as-a-Judge Human-like Summarization Evaluation with ChatGPT

Reference 2006

Resolution
unresolved
no resolver link, observed 2026-08-05T22:40:42.192331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:40:42.192331Z digest=sha256:a142d62aeb1f795e3cccd0726f64e01b47e49eb3939bcf50d6c14219e22a3e4f

Observation 136b11e9-e3d3-4a0a-81a7-45a6755596d3 · inbound

Towards Personalized Explanations for Health Simulations: A Mixed-Methods Framework for Stakeholder-Centric Summarization cites this paper.

Towards Personalized Explanations for Health Simulations: A Mixed-Methods Framework for Stakeholder-Centric Summarization Human-like Summarization Evaluation with ChatGPT

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T05:59:50.428652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:59:50.428652Z digest=sha256:32c6967ebecdf2b1e5e9fca474a07bf0abbc6bf37772a473b4ae720b9387b910

Observation 9cf5c829-93f6-42de-be24-672595c7a65e · inbound

AraHalluEval: A Fine-grained Hallucination Evaluation Framework for Arabic LLMs cites this paper.

AraHalluEval: A Fine-grained Hallucination Evaluation Framework for Arabic LLMs Human-like Summarization Evaluation with ChatGPT

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:05.221826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T06:00:05.221826Z digest=sha256:f64610c28cb37e6fe1cb4c95a2c289600c68b51a36c6727d12730c9f2a345290

Observation 5e33f5bb-19b6-4662-ab56-48f6373323aa · inbound

Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning cites this paper.

Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning Human-like Summarization Evaluation with ChatGPT

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T14:49:51.562656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:49:51.562656Z digest=sha256:2af2635372a0a7f202d71f71087835dd29ac96fbe634c2d70415f68e4d60dcbd

Observation fc0e4375-8325-4003-854b-0831ba72c975 · inbound

From Moderation to Mediation: Can LLMs Serve as Mediators in Online Flame Wars? cites this paper.

From Moderation to Mediation: Can LLMs Serve as Mediators in Online Flame Wars? Human-like Summarization Evaluation with ChatGPT

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T18:57:00.265923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:57:00.265923Z digest=sha256:725e1520162fa1021d1d7089402e2556976a070068426d44d1721c3a35f694c9

Observation 98df71dd-262c-4bc6-a17e-7cb5936d0a21 · inbound

User Perceptions of an LLM-Based Chatbot for Cognitive Reappraisal of Stress: Feasibility Study cites this paper.

User Perceptions of an LLM-Based Chatbot for Cognitive Reappraisal of Stress: Feasibility Study Human-like Summarization Evaluation with ChatGPT

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T13:05:28.935067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:05:28.935067Z digest=sha256:5e9ee4b68aef59ca540f8243bb5ff9e9e72f7c300c930a941560b88847f07525

Observation 704a3560-05c3-4bcc-87d5-054ee716e1fc · inbound

Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory cites this paper.

Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory Human-like Summarization Evaluation with ChatGPT

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T06:08:26.400539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:08:26.400539Z digest=sha256:aa6b4e72df82be2a4811db23bfdbede01f582fd746c30482a81bb3affd66a116

Observation 0dfd8e45-4b92-4d8b-b3ce-17fc499899f5 · inbound

LLM-ReSum: A Framework for LLM Reflective Summarization through Self-Evaluation cites this paper.

LLM-ReSum: A Framework for LLM Reflective Summarization through Self-Evaluation Human-like Summarization Evaluation with ChatGPT

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:46:37.680608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T16:20:50.111108Z digest=sha256:5f7de533e794be9e2561c5e9d408af580d3f61a35a0270f607b8356a6906a857

Observation 7c59b5e1-23bc-45bf-9f81-85e651965119 · inbound

Bridging Reasoning Trajectories in On-Policy Distillation via Near-Future Guidance cites this paper.

Bridging Reasoning Trajectories in On-Policy Distillation via Near-Future Guidance Human-like Summarization Evaluation with ChatGPT

Reference 104

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T10:44:37.284099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-30T10:34:52.474916Z digest=sha256:acc335ed6f2ececb7bfa9c5ac3cfac59c6a963198e9f753361678bc8c66d6e40

Observation 21bfc490-1d5f-464a-8b0a-c7a871e1b691 · inbound

Illusions of the Gold Standard: A Large-scale Analysis of Human Evaluation Protocols for Long-form Text Generation cites this paper.

Illusions of the Gold Standard: A Large-scale Analysis of Human Evaluation Protocols for Long-form Text Generation Human-like Summarization Evaluation with ChatGPT

Reference 114

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:37:22.701785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T20:20:08.996005Z digest=sha256:33b53da314c706290499b90262fc629651c9a6592c51e8c5533e9cd16e53f99b