Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2406.11044.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:34:00.073227Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-23T17:35:43.963141Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation d7b7f57a-d576-432d-833c-6f084ab4d494 · inbound
A Survey on LLM-as-a-Judge Evaluating the Performance of Large Language Models via Debates
Reference 109
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4920742b-8605-4fce-b84a-a72a01b7563d · inbound
LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods Evaluating the Performance of Large Language Models via Debates
Reference 163
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b945e715-2bc8-4460-bcaa-b799810c3a82 · inbound
LLM-based Human Simulations Have Not Yet Been Reliable Evaluating the Performance of Large Language Models via Debates
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d338e07-2382-4d6e-895e-b3229c4c8e43 · inbound
The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs Evaluating the Performance of Large Language Models via Debates
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78246a18-2786-45dd-9476-49eb94c52a75 · inbound
Efficient MAP Estimation of LLM Judgment Performance with Prior Transfer Evaluating the Performance of Large Language Models via Debates
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e966dac2-cd50-4d15-baaa-94f3454d0b18 · inbound
Pretraining on the Test Set Is No Longer All You Need: A Debate-Driven Approach to QA Benchmarks Evaluating the Performance of Large Language Models via Debates
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.