Pith. sign in

Paper Citation Record · LEDGER

LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2307.07889.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.07889 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:21:26.022925Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

5
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 35823cd7-a060-4ab9-aea8-d1ab449fe8db · inbound

An Empirical Study of Evaluating Long-form Question Answering cites this paper.

An Empirical Study of Evaluating Long-form Question Answering LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T10:21:26.022925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:21:26.022925Z digest=sha256:ae190b078f805e0b412e2d3d05a4d4fa8bccaae35492f653a6143a2a38f5faa1

Observation 276b1c53-213e-435c-a784-07fa241da9f5 · inbound

MSQA: Benchmarking LLMs on Graduate-Level Materials Science Reasoning and Knowledge cites this paper.

MSQA: Benchmarking LLMs on Graduate-Level Materials Science Reasoning and Knowledge LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T12:43:28.289135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:43:28.289135Z digest=sha256:52d45ae0154843b6d0f4796167d190342ab4f06c9447e6bd74930966b82df5ea

Observation df6d0917-afb7-454d-b77f-f2a6d1a7b858 · inbound

Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons cites this paper.

Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T10:58:27.025105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:58:27.025105Z digest=sha256:f0c31a08b0b737947f53b4b8fa814df95602550f4337e23c50a3c88b7a6236bb

Observation 9293ed8f-8396-4bea-a513-a27f94704542 · inbound

BiT-MCTS: A Theme-based Bidirectional MCTS Approach to Chinese Fiction Generation cites this paper.

BiT-MCTS: A Theme-based Bidirectional MCTS Approach to Chinese Fiction Generation LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-15T11:25:30.953783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T11:22:36.493502Z digest=sha256:59dc5a70331b9648ec5fd4982ca82f6e032034cf7a4ed0474f2420c91c4d9cb4

Observation a3cee2e0-2d60-4e12-8634-d21aeba4612d · inbound

Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring cites this paper.

Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:51:03.660848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-10T04:31:53.825854Z digest=sha256:052a4b4cbc7644173ae02d2f2ee20c416dd57bcab6824793b71bc9de309defda

Observation 2d92ae1e-f0f4-402d-bd26-fff15ae986f0 · inbound

MIRAI: Prediction and Generation of High-Impact Academic Research cites this paper.

MIRAI: Prediction and Generation of High-Impact Academic Research LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:06:55.975498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T02:29:38.654022Z digest=sha256:49f7baf5d004639c6d985d1fbde867a816b6a1d028bd288ade213500bfba7a17

Observation 42299ee0-80e4-49aa-a0c9-8965dd59457d · inbound

LLM-Based Examination of Eligibility Criteria from Securities Prospectuses at the German Central Bank cites this paper.

LLM-Based Examination of Eligibility Criteria from Securities Prospectuses at the German Central Bank LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-26T03:58:57.097881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-26T03:56:28.271760Z digest=sha256:b5efdc689366f877091c40cb9296ee06ed3c305d2f30bdb6ec30e36862cbd28c

Observation 7fcaceed-86de-4e1a-8ae4-445f69edd414 · inbound

Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026 cites this paper.

Causal Connections: Leveraging Multilingual Fine-Tuning for Financial QA@FinCausal 2026 LLM Comparative Assessment: Zero-shot NLG Evaluation through Pairwise Comparisons using Large Language Models

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-06-29T02:23:00.947061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-29T02:17:26.872854Z digest=sha256:a92e6dd5de2a0ca83d314d3cd751a40a8e1eb142112f74e1ed41fcab42a2b6d4