Pith. sign in

Paper Citation Record · LEDGER

Pairwise or Pointwise? Evaluating Feedback Protocols for Bias in LLM-Based Evaluation

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2504.14716.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.14716 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T03:55:51.871628Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5723ceee-0c3e-483e-a676-cedb5842377e · inbound

FairJudge: An Adaptive, Debiased, and Consistent LLM-as-a-Judge cites this paper.

FairJudge: An Adaptive, Debiased, and Consistent LLM-as-a-Judge Pairwise or Pointwise? Evaluating Feedback Protocols for Bias in LLM-Based Evaluation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T03:55:51.871628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:55:51.871628Z digest=sha256:de4b572d353e6ce72deb8c55ddf3e4b244658cf3d3733c5d5c7aa02f8bc93d33

Observation 15b55d64-3d15-46d6-b1c8-658d493bac4b · inbound

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning cites this paper.

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning Pairwise or Pointwise? Evaluating Feedback Protocols for Bias in LLM-Based Evaluation

Reference 112

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:15:48.815431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T19:15:27.406778Z digest=sha256:2f83783a984075d3940dfe5bf1349df7cec7f72c1103f772ec2f32bf3d206d85

Observation 0b965990-9db0-4168-a850-c99fae2e99fc · inbound

JudgmentBench: Comparing Rubric and Preference Evaluation for Quality Assessment cites this paper.

JudgmentBench: Comparing Rubric and Preference Evaluation for Quality Assessment Pairwise or Pointwise? Evaluating Feedback Protocols for Bias in LLM-Based Evaluation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:24:37.712899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T11:16:19.162691Z digest=sha256:e9f9d98632fb049e6528432ceb3c8f6685bd36d41593761c24b8b5ee890bdcfa

Observation 5d77a0b2-3371-4c41-896f-4ee1da1fad26 · inbound

Trust Region On-Policy Distillation cites this paper.

Trust Region On-Policy Distillation Pairwise or Pointwise? Evaluating Feedback Protocols for Bias in LLM-Based Evaluation

Reference 152

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T20:56:13.559273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T17:38:50.313305Z digest=sha256:bfab3031bcf4e3305b8e973f710e9bca6a8536c0b2bc4fa774c6c5abac5bc0d6

Observation fb1fe938-d861-44d0-81b7-ca8eee5a7336 · inbound

Towards Fast Domain Adaptation and Fine-Grained User Simulation for Evaluating Conversational Recommender Systems cites this paper.

Towards Fast Domain Adaptation and Fine-Grained User Simulation for Evaluating Conversational Recommender Systems Pairwise or Pointwise? Evaluating Feedback Protocols for Bias in LLM-Based Evaluation

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:09:48.841413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T07:13:21.198407Z digest=sha256:618dda9a3c2aaa129b4a007b4aa5cc098c6e0f95bd64581d59b3d25c9badf71a