Pith. sign in

Paper Citation Record · LEDGER

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat

As of 19 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 2 inbound Pith citation observations for arXiv:2411.14483.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14483 v2

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T17:14:07.204201Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:51:45.801126Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T15:12:29.896082Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0faa153b-8cc9-45c1-9e35-cacb099a0f87 · outbound

This paper cites online" 'onlinestring :=.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.113437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.113437Z digest=sha256:7284c86ea860c452ed4d595369fe0e9991515bcf42436a55c7cf1a48c911cbe4

Observation 5ea6b7f8-37d7-4f15-af7e-794d9dd9675a · outbound

This paper cites write newline.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.118103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.118103Z digest=sha256:57f872577f7cd2d47dd3906f5b2a60d5d89d1051d48c2fb43185fc6421628b79

Observation d1df92e1-a4fc-4fb0-8af5-f59bff7d72a2 · outbound

This paper cites Elo Uncovered: Robustness and Best Practices in Language Model Evaluation.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Elo Uncovered: Robustness and Best Practices in Language Model Evaluation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.122126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.122126Z digest=sha256:85307b398cfc39af114bf71c64e3cdd20b004c4d25a8ce8f59337158837b5db7

Observation 0fc82042-0240-439b-83a4-0c6d43bbdfec · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.126918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.126918Z digest=sha256:4c8515d661a516d5c0256e5a17db74d255af98ec21d90b08d181bfbb7239dd18

Observation 2e2c5c88-ce1b-433f-8b35-56f1a4afb37c · outbound

This paper cites Random Walker Ranking for NCAA Division I-A Football.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Random Walker Ranking for NCAA Division I-A Football

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-12T17:14:07.280389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.130593Z digest=sha256:dbec8db046c422f16d764ec7f82f8f0f8a56ba1cfe5340bbce7e8267fda4e3e1

Observation 8745f801-9476-442d-a55e-36dc78b4e8f5 · outbound

This paper cites The Bowl Championship Series: A Mathematical Review.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat The Bowl Championship Series: A Mathematical Review

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-12T17:14:07.262409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.135694Z digest=sha256:657d27512071bff404160c6b225741bdc0a9d03606460f1ad679e4a3ae1796cc

Observation 7b4ccd4e-50ab-4fc4-a132-2de46853ee93 · outbound

This paper cites Gonzalez, Ion Stoica, and Eric P.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Gonzalez, Ion Stoica, and Eric P

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.141008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.141008Z digest=sha256:37efb8edef44a38ccc46076ff53bcbe9a043c37d8feb1a3921f2b149042b845d

Observation f82815bd-a9a5-4175-9ab3-d4f106ab3ba9 · outbound

This paper cites Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.145202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.145202Z digest=sha256:972db4cef142ffadb09926340a5f507b7dd165a667b9daac6b37b2a8be557588

Observation 77e2b03e-967c-4894-b92e-4007a3548ecd · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:14:07.434268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.149523Z digest=sha256:dcef07bd18dac8e5c445301151a612d3c6843defe94f81853eada81ddd6a1182

Observation 88da050e-756d-4203-bc16-765af77b21f9 · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.153215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.153215Z digest=sha256:cda9851da2ce1132807a296dda4b9911e15e1d8650b7d2aa97426b06a086e6e9

Observation 901f8317-a6d8-4037-aa23-4b3158a60b7a · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:14:07.410615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.157474Z digest=sha256:aba95a43bea51a902244d4e989b0776c6766bb2ceea84cc6390b4cd5e7c20699

Observation 31a274a7-8fc9-4333-a1ee-795c28588d9f · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.161722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.161722Z digest=sha256:f9a3a28fc4b4292f908ad9383ab34a2b9f1a364c30b77613635a38a331b6ca8e

Observation 69201b8c-3014-4676-add3-3fd8a804dba2 · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:14:07.398399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.165801Z digest=sha256:901abfa21ec77fa8d5cf2b0006ddf1083b9d2958ae851d80ec5c0b011df4c261

Observation 86b0d2e0-f74a-4db8-be1e-aa9a95e20b1f · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:14:07.386661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.170154Z digest=sha256:0c5e43f6bfc65a75147c398202d9721e58bd3c7d14a17424663e67a03ce3bc21

Observation 97604a30-f602-487b-84c3-c88f0e8f2cff · outbound

This paper cites Scaling Down to Scale Up: A Cost-Benefit Analysis of Replacing OpenAI's LLM with Open Source SLMs in Production.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Scaling Down to Scale Up: A Cost-Benefit Analysis of Replacing OpenAI's LLM with Open Source SLMs in Production

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.173862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.173862Z digest=sha256:94d59b79a9066771b96bbc78e5f29cee5df27972f44c004e692a9ff1e4a78fe3

Observation dfc8f062-b4b0-49e1-8db4-a2e7280e0138 · outbound

This paper cites LLM-Blender: Ensembling Large Language Models with Pairwise Ranking and Generative Fusion.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat LLM-Blender: Ensembling Large Language Models with Pairwise Ranking and Generative Fusion

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.177986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.177986Z digest=sha256:5559f9cc3366799cbcea2fd8419ea4beac080834c619452024ede277b8bc2013

Observation 5d87537c-57e1-4a35-b2e2-ea76946e0077 · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:14:07.374231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.183153Z digest=sha256:34a436faa289e17e0606b6a1433b3f3f5eab542e5e32e33189c139095daf7c80

Observation 48450322-31bb-4ba7-8003-20c09d4383e5 · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:14:07.361647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T17:14:07.187372Z digest=sha256:827366ca3b5b3fa15b9cc144c88b44ac0a3e16d636457af564d46fe7c3ae3ab8

Observation 8502a7b6-ba61-47de-ba23-4cf32a818f38 · outbound

This paper cites SuperGLUE: A Stickier Benchmark for General-Purpose Language Understanding Systems.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat SuperGLUE: A Stickier Benchmark for General-Purpose Language Understanding Systems

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.191359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.191359Z digest=sha256:3a321d245aeb798ad1b86df01135835abf376d5f2288a3f294fbf88ba233f3d3

Observation eadf0244-935b-4c52-93b3-eb4560620dd3 · outbound

This paper cites GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.195683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.195683Z digest=sha256:1df35393654270b5908e1c0da8f0a13dee7d82744e19efadcc33fa4d61d9ccce

Observation 22f145d0-e36f-48f7-999d-d95152ca568a · outbound

This paper cites Style Over Substance: Evaluation Biases for Large Language Models.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Style Over Substance: Evaluation Biases for Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.199865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.199865Z digest=sha256:c45bd83a3a18d221f87d1a29efb67fb64e10eaefb552fa1569e4fbe101f51c81

Observation 32e177b6-3461-4d0b-8a85-195111fd4784 · outbound

This paper cites an unresolved cited work.

Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T17:14:07.204201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:14:07.204201Z digest=sha256:21004ca7a4836400591b034ed86be7ac83d1a5720987a67dad4f5335fdba6a5d

Pith citing papers

Observation 50be9836-8e8d-4c41-a469-89c5026fa4b9 · inbound

Re-evaluating Automatic LLM System Ranking for Alignment with Human Preference cites this paper.

Re-evaluating Automatic LLM System Ranking for Alignment with Human Preference Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat

Reference 1952

Resolution
unresolved
no resolver link, observed 2026-08-10T22:51:45.801126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:51:45.801126Z digest=sha256:8fc5e3ecd923afe3eb4063395ae6e577b0a815bee8237e76f7d2f95b88ff9b51

Observation c7de3eed-8795-47f5-ba78-8065e605404f · inbound

SLMEval: Entropy-Based Calibration for Human-Aligned Evaluation of Large Language Models cites this paper.

SLMEval: Entropy-Based Calibration for Human-Aligned Evaluation of Large Language Models Ranking Unraveled: Recipes for LLM Rankings in Head-to-Head AI Combat

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:12:29.990097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T15:12:26.391373Z digest=sha256:de2737255d66a053e21c36829521dac8d82873fe01e9c775e0f0f2f10fe46775