Pith. sign in

Paper Citation Record · LEDGER

A Closer Look into Automatic Evaluation Using Large Language Models

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2310.05657.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.05657 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:56:44.035946Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T13:21:24.424456Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cff20c7d-4cb6-46c8-a067-effaa1d92336 · inbound

DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models cites this paper.

DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models A Closer Look into Automatic Evaluation Using Large Language Models

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-16T13:21:24.426800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-16T13:21:24.297836Z digest=sha256:6a17a051b997bae8b46a2893e908e84efad913eec6f4b55ecb9831fce3948799

Observation e163aa16-5511-4ac9-a9ee-7262dde01c41 · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods A Closer Look into Automatic Evaluation Using Large Language Models

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:08:36.213435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:55c7506398f673c57f306f881d02b4632e8ac5dd701bd799138ebb3bc3f87c47

Observation 52f0b447-b60c-45b0-9fb7-fff4ca4b4a4a · inbound

FECT: Factuality Evaluation of Interpretive AI-Generated Claims in Contact Center Conversation Transcripts cites this paper.

FECT: Factuality Evaluation of Interpretive AI-Generated Claims in Contact Center Conversation Transcripts A Closer Look into Automatic Evaluation Using Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T13:56:44.035946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:56:44.035946Z digest=sha256:567291ba4671340859c078613da0835fedcf92fa558a6bc81b2f06daff9cafcc

Observation d63e359a-b106-4ba5-863b-6b3ffcb2ccbe · inbound

PEEM: Prompt Engineering Evaluation Metrics for Interpretable Joint Evaluation of Prompts and Responses cites this paper.

PEEM: Prompt Engineering Evaluation Metrics for Interpretable Joint Evaluation of Prompts and Responses A Closer Look into Automatic Evaluation Using Large Language Models

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T14:00:02.948664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T13:57:41.428695Z digest=sha256:42e6b08b248cbc7305754f06cfd6cb4a18dff25650cd0d87ad0ddc7e350be65f

Observation ad93f109-7739-4ce9-9fde-edd5fa5eda8e · inbound

Towards Annotation-Free Validation of MLLMs: A Vision-Language Logical Consistency Metric cites this paper.

Towards Annotation-Free Validation of MLLMs: A Vision-Language Logical Consistency Metric A Closer Look into Automatic Evaluation Using Large Language Models

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:06:11.378608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T10:18:37.901403Z digest=sha256:5fba7a209f30dfda26b7da4d02a2013ef1b7c6fe3bf2638188d9fe34ce68c4d9