Pith. sign in

Paper Citation Record · LEDGER

Shepherd: A Critic for Language Model Generation

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2308.04592.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.04592 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:12:36.352474Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T17:35:43.992836Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation be7f4ce0-1816-411e-b669-6da173afed77 · inbound

Large Language Models Cannot Self-Correct Reasoning Yet cites this paper.

Large Language Models Cannot Self-Correct Reasoning Yet Shepherd: A Critic for Language Model Generation

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:48:27.119090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T05:48:24.133045Z digest=sha256:801841124659c50316bccf7f3c331e819c88f9417d390b28a2681f2038b95cd0

Observation 95f367e4-9055-43fa-b6d2-7909ae475abf · inbound

A Survey on LLM-as-a-Judge cites this paper.

A Survey on LLM-as-a-Judge Shepherd: A Critic for Language Model Generation

Reference 163

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:35:43.995480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T17:33:13.394338Z digest=sha256:1aef245df545bf910518b5db0ee056bbcdeb51d7b7e18d95f32d42ea0f9d14a4

Observation 785ceb04-4501-4828-9b9c-0298c0b1ab02 · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods Shepherd: A Critic for Language Model Generation

Reference 241

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:08:35.631356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:5c3e12a5ae3bee30a156a45c59ef1914ec55553c0401d1297942321f000559f6

Observation 830c0775-43ec-4910-8ac5-6d22604b3beb · inbound

HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs cites this paper.

HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs Shepherd: A Critic for Language Model Generation

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:36:50.215433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T12:36:50.060335Z digest=sha256:d49f75d191b3ea783054894a7ae30c6581903efd909516994115bacf0a2d01d1

Observation 1d23d641-c885-4c6d-9156-b90c64e84549 · inbound

Reinforcement Learning from Human Feedback cites this paper.

Reinforcement Learning from Human Feedback Shepherd: A Critic for Language Model Generation

Reference 293

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:32:01.170999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T19:27:40.991325Z digest=sha256:57fef5886bba320aa7b0b16fae21cda0068d87e766b2f4593386e075424b1d4c

Observation 56ae71b1-4c32-4e6b-bd47-4094ef1b9bac · inbound

SkillVerse : Assessing and Enhancing LLMs with Tree Evaluation cites this paper.

SkillVerse : Assessing and Enhancing LLMs with Tree Evaluation Shepherd: A Critic for Language Model Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:36.352474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:12:36.352474Z digest=sha256:70bab69489e7338537eed6fe0fbb0d8f57f5002d8229a66a7c18d840eecb8dfc

Observation 98891d74-f711-49f2-b972-8e3616af3e5c · inbound

To trust or not to trust: Attention-based Trust Management for LLM Multi-Agent Systems cites this paper.

To trust or not to trust: Attention-based Trust Management for LLM Multi-Agent Systems Shepherd: A Critic for Language Model Generation

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:32:17.409057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T11:30:47.877793Z digest=sha256:273d4c0e1dd980f4aab0553327227612f301bbc32894fd9cb907c6a74b561c28

Observation da1e0fd0-387e-432e-ae69-46dc55fc6705 · inbound

Exchange of Perspective Prompting Enhances Reasoning in Large Language Models cites this paper.

Exchange of Perspective Prompting Enhances Reasoning in Large Language Models Shepherd: A Critic for Language Model Generation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:00.133441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:00.133441Z digest=sha256:57ed48c713b76ca52c74f3927e6a7a73816a614482c11024cb400b5288986f3d

Observation 43ef97b8-84fb-400b-b6f3-4abebaaa0c3e · inbound

Understanding Human Limits in Pattern Recognition: A Computational Model of Sequential Reasoning in Rock, Paper, Scissors cites this paper.

Understanding Human Limits in Pattern Recognition: A Computational Model of Sequential Reasoning in Rock, Paper, Scissors Shepherd: A Critic for Language Model Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T14:26:03.582133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:26:03.582133Z digest=sha256:e0e7c566f53d439c96e9f3ebebfd679260ed6b838fa44a50943e389c775db1a7

Observation cb10a7f6-4dd5-403f-b8af-675e9a7cf402 · inbound

User-Assistant Bias in LLMs cites this paper.

User-Assistant Bias in LLMs Shepherd: A Critic for Language Model Generation

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T22:51:53.108335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T22:51:02.400926Z digest=sha256:80e28f783a4d9eed0a6ae8215dd5910ae9eb05eafba22fc967658a6d2caa0780

Observation 61b259b0-be66-419e-8f3f-b5a29f70e516 · inbound

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination cites this paper.

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination Shepherd: A Critic for Language Model Generation

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T18:02:42.359251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-19T18:01:06.649723Z digest=sha256:b23ac61e60c3af7eaab5d95e8ef9532d212c8030976a7efcc743e9be693e3d0e

Observation b7d42c34-5f16-46a9-9cb9-82521729b748 · inbound

Meta-Learned Reward Shaping for Reinforcement Learning from Human Feedback cites this paper.

Meta-Learned Reward Shaping for Reinforcement Learning from Human Feedback Shepherd: A Critic for Language Model Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T03:14:12.848281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:14:12.848281Z digest=sha256:e7022be270d248bac7e0390b56a5c520d34178af36bd307a2d4ac9d02e3d8ef0