Pith. sign in

Paper Citation Record · LEDGER

Rethinking Human Preference Evaluation of LLM Rationales

As of 9 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 1 inbound Pith citation observation for arXiv:2509.11026.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.11026 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T17:15:27.898978Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T12:47:37.888316Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 178274f0-4833-4e24-b46d-ebaa34776a55 · outbound

This paper cites REV: Information-Theoretic Evaluation of Free-Text Rationales.

Rethinking Human Preference Evaluation of LLM Rationales REV: Information-Theoretic Evaluation of Free-Text Rationales

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.861042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.861042Z digest=sha256:2e8d730939e574d96d5c741eed8d25c6522179b908e60eb477cf96b0beae4638

Observation 05322b96-4ea8-4209-8bb7-3085c46d7962 · outbound

This paper cites Deep reinforcement learning from human preferences.

Rethinking Human Preference Evaluation of LLM Rationales Deep reinforcement learning from human preferences

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.866035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.866035Z digest=sha256:f8d5c2ff8a68ccbb84325f2a09d45cf15da54f9552e8826eef694e03e57cb3da

Observation 4a6dd9a8-749d-476d-b2fd-8ef1e599dc4a · outbound

This paper cites ROSCOE: A Suite of Metrics for Scoring Step-by-Step Reasoning.

Rethinking Human Preference Evaluation of LLM Rationales ROSCOE: A Suite of Metrics for Scoring Step-by-Step Reasoning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.868447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.868447Z digest=sha256:bb5cf193cafbf11f4d3eb191a6b5c55d78a6e38d7da10bbe6f962e8e1e0a84f2

Observation 71f6839e-b693-4227-8801-8a1b6fb6d790 · outbound

This paper cites Are Machine Rationales (Not) Useful to Humans? Measuring and Improving Human Utility of Free-Text Rationales.

Rethinking Human Preference Evaluation of LLM Rationales Are Machine Rationales (Not) Useful to Humans? Measuring and Improving Human Utility of Free-Text Rationales

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.875563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.875563Z digest=sha256:20ebd4df4f6e4bad052d6fbeeb881b509b0973d82dcd01c5918d75fcf9926230

Observation 2996221c-3aa0-4582-8b8d-822dc2b55ca9 · outbound

This paper cites 2 OLMo 2 Furious.

Rethinking Human Preference Evaluation of LLM Rationales 2 OLMo 2 Furious

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.882256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.882256Z digest=sha256:4eebd16dd2cd068e81a3ea820943995ef2db3e8a50a2d0edac5b138b3a37bcae

Observation d60bde18-2647-4359-b013-92f3d56a5127 · outbound

This paper cites Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al.

Rethinking Human Preference Evaluation of LLM Rationales Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.884639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.884639Z digest=sha256:f0ecbd433652b6561e12a42329df18f8614ffdb8e6bca4f3f65686ad9e89e18f

Observation b8e998bf-7eae-4b59-bb1d-43c97d681ccf · outbound

This paper cites ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness.

Rethinking Human Preference Evaluation of LLM Rationales ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.887155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.887155Z digest=sha256:e6b3f647338a1bd02804a3d1cd4a4b79eb5ac6a209a9fbca34c419321da47263

Observation ea864c49-05f0-41b0-bb51-7c3494cd308c · outbound

This paper cites Tailoring Self-Rationalizers with Multi-Reward Distillation.

Rethinking Human Preference Evaluation of LLM Rationales Tailoring Self-Rationalizers with Multi-Reward Distillation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.891750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.891750Z digest=sha256:c4e367f88633a26a9857c5bf6f3a66a5200dcfe3026ff91d7242b4f1262f0429

Observation fd8281ac-fab5-489a-81d8-c777d494e750 · outbound

This paper cites PINTO: Faithful Language Reasoning Using Prompt-Generated Rationales.

Rethinking Human Preference Evaluation of LLM Rationales PINTO: Faithful Language Reasoning Using Prompt-Generated Rationales

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.894185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.894185Z digest=sha256:ac1ac6478a49c6a88fa4e168e1c6c1c520d558a15213de8103ea02fdb0bead70

Observation 9a58348b-4a58-4cc3-ad89-25bfcc4479f0 · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

Rethinking Human Preference Evaluation of LLM Rationales Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.898978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.898978Z digest=sha256:66413c611d8a6287c8dc37d8e39bd4ae658c1833200126aaf4dec85e7de37fcd

Observation 3858e017-d4a9-4b83-8191-6be19cb770b7 · outbound

This paper cites cc/paper files/paper/2016/file/10a5ab2db37feedfdeaab192ead4ac0e-Paper.pdf.

Rethinking Human Preference Evaluation of LLM Rationales cc/paper files/paper/2016/file/10a5ab2db37feedfdeaab192ead4ac0e-Paper.pdf

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.880107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.880107Z digest=sha256:b278df392b2b38f15015d732fc87aae6f4566ad55b74d3fc769ab4b638444ac3

Observation 1e27b531-77ab-4288-b346-651276ab82ed · outbound

This paper cites Scott M Lundberg and Su-In Lee.

Rethinking Human Preference Evaluation of LLM Rationales Scott M Lundberg and Su-In Lee

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.877843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.877843Z digest=sha256:f02a8ec84fc963277e034fde8cdc15b4c74afe3f84b5507e20b5955d8f8ccecf

Observation 5de3ddad-9dde-4a3e-ab0f-98e9b50cd13a · outbound

This paper cites Explain Yourself! Leveraging Language Models for Commonsense Reasoning.

Rethinking Human Preference Evaluation of LLM Rationales Explain Yourself! Leveraging Language Models for Commonsense Reasoning

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.889372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.889372Z digest=sha256:d358be01350409c3f5918934c184249fca6c6d868994d782e80710f16dc0b4bb

Observation 3243e181-b116-4136-88f6-dbe0913ea7a6 · outbound

This paper cites DecipherPref: Analyzing Influential Factors in Human Preference Judgments via GPT-4.

Rethinking Human Preference Evaluation of LLM Rationales DecipherPref: Analyzing Influential Factors in Human Preference Judgments via GPT-4

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.873148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.873148Z digest=sha256:9b3ef488fc505b1b727e31513ae9f1a5daa555fb32af3955601c1860754cc5a1

Observation f740dcd0-d594-4249-84bb-70da27720075 · outbound

This paper cites Reframing Human-AI Collaboration for Generating Free-Text Explanations.

Rethinking Human Preference Evaluation of LLM Rationales Reframing Human-AI Collaboration for Generating Free-Text Explanations

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.896357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.896357Z digest=sha256:ef975ec46edb3bc7c5a3b89d064d24c40a10cf39b873675c7840caa828cdad21

Observation c877e8ca-d434-4b45-a792-6b23e27f6ff8 · outbound

This paper cites Faithfulness Tests for Natural Language Explanations.

Rethinking Human Preference Evaluation of LLM Rationales Faithfulness Tests for Natural Language Explanations

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.857734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.857734Z digest=sha256:ac26ec55017c0574d7d16d7d00a7c3cef470fd2b7cbc842683bba4c8a6b50cf0

Observation b9c9a3ff-c413-4d5f-845b-cc6fadf33474 · outbound

This paper cites Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference.

Rethinking Human Preference Evaluation of LLM Rationales Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.863520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.863520Z digest=sha256:884340eba9053beee1738fa6612f07bc4c3fb881d7445a002750eb569b364864

Observation 1b80c55b-79d6-4cf8-ac00-58ff29db57d8 · outbound

This paper cites an unresolved cited work.

Rethinking Human Preference Evaluation of LLM Rationales Unresolved cited work

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T17:15:27.870816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:15:27.870816Z digest=sha256:1d909c93a3063596f26f43c7ee1eb07a514c0ccca526f8f6869133364def8dcb

Pith citing papers

Observation 50e6b86a-e429-4e72-a484-99f4ea2e78e5 · inbound

CodeRescue: Budget-Calibrated Recovery Routing for Coding Agents cites this paper.

CodeRescue: Budget-Calibrated Recovery Routing for Coding Agents Rethinking Human Preference Evaluation of LLM Rationales

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T12:47:37.888316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:47:37.888316Z digest=sha256:67fe37b38ff5c2ce6049c597aa438de1c99bc48029652a193a6744e649b090e8