Pith. sign in

Paper Citation Record · LEDGER

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability

As of 21 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2504.13972.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.13972 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:15:28.429441Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact6
  • verified fuzzy0
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 25fa1dfd-a10d-472b-8bd0-01288751c2fb · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:15:28.983959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.306729Z digest=sha256:b2bab0847d330b14efe302d76af0fb4ace54bfda706220bf2c8c572841ac1c12

Observation ec6a8a45-0f9e-4f02-89ca-c02fa7921577 · outbound

This paper cites Language Models (Mostly) Know What They Know.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Language Models (Mostly) Know What They Know

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T12:15:28.311892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:15:28.311892Z digest=sha256:d6ba9ec5706e7404e98ffddc7cb0dca721c4edaa4e2b9b626844cc13d40ff349

Observation 661ea523-5de7-4b81-b424-a7554c7af86b · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:15:28.968686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.317289Z digest=sha256:64d8133acc07a5932c3e6c6f0669405e201cb8307777b5e9629aca1bf5734806

Observation fd492477-2f6c-438a-a44f-63d855a7076e · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T12:15:28.322279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:15:28.322279Z digest=sha256:88ccb5ad3fc29025226067bc7bcbff72c44d717a05ec694040e239ea5e397a37

Observation ec23ae42-f94c-42b1-9e31-caaacc94eff2 · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:15:28.943413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.327146Z digest=sha256:8699b0fa69c4a66edc3ba34dc1b921378d0d3d8e8b068789c704ed2d9d19dd70

Observation 254ae448-07a8-477e-af5c-1b02bf2cbe0a · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:15:28.927976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.332298Z digest=sha256:04feb5fe609dbd5f3014fd6b87f395e550f3a1717691d06dcbff31c439f88169

Observation 13ceab8a-f8ab-4ead-8b76-24cb0dc09bf8 · outbound

This paper cites Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T12:15:28.337487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:15:28.337487Z digest=sha256:7795d7a030ee13028e2fc5a1d98a759152a11d57d5ea7ac77cf31fdda91a356a

Observation 9acb53bd-6081-4b2d-9885-527d5cc8afc6 · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:15:28.912740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.342484Z digest=sha256:b9916135b6076825ba93a3231e13790c884fcf62004250f12f7f6592684ac469

Observation 86253027-95dc-4d81-a38b-72c3c448911b · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:15:28.898603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.347007Z digest=sha256:212b2df39a7eeddb97978e92ecc0305cbb789b1242b923d8920acb4c746fd0a6

Observation 42a3de23-109b-4832-8f3d-12f01c8d64f4 · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:15:28.884696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.351721Z digest=sha256:f582f6a1b8c0b6ba672337fab33b68c0b4779017c3f47521ccd1a2dc67541f10

Observation 0dba8b77-90d7-4e4c-9a71-fe50c58975dc · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:15:28.870918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.356268Z digest=sha256:695948963f6d523e4414c667be2afe19c679ab4ece18d32011e128dca6e13f43

Observation 4b588c66-a11e-4e20-86b8-06a1e1117f42 · outbound

This paper cites Having Beer after Prayer? Measuring Cultural Bias in Large Language Models.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Having Beer after Prayer? Measuring Cultural Bias in Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T12:15:28.360877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:15:28.360877Z digest=sha256:3549500deb75805a4cf7af8eefaa39bb4f078a6f9b120b167015f3149529dbb0

Observation c67ba1f3-03ea-4922-b84b-d0bcc8628015 · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:15:28.856509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.366328Z digest=sha256:c6baeb9ed07a297be1258873d15e1f9c944eab578d845accde18431a0fdbff7e

Observation 74e69db7-fc30-4bcb-8d05-849e44bb8b1d · outbound

This paper cites GPT-4 Technical Report.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability GPT-4 Technical Report

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T12:15:28.370960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:15:28.370960Z digest=sha256:fa0209e4bf1f41cead17064d184744ae99b930b5054eaec417209e44e5c29398

Observation aa3abb44-d74a-42a3-a67b-4e3e2d120841 · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T12:15:28.375473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:15:28.375473Z digest=sha256:c9cd7a1f2dd92b3e06297a02a0b78b85ca660856f57fc85d6689d8ef0c75864e

Observation d5333458-f460-4f83-a574-12df1db15104 · outbound

This paper cites Internet Service Providers' and Individuals' Attitudes, Barriers, and Incentives to Secure IoT.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Internet Service Providers' and Individuals' Attitudes, Barriers, and Incentives to Secure IoT

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-16T12:15:28.732058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.379964Z digest=sha256:db9e92be35b957a5612d0004c10e11bbcd73b1ea5b95336308a28729aff6d204

Observation f8987cdd-71e6-4efc-83c1-5c05a28fcbb7 · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 17

Resolution
verified exact
raw_fallback, observed 2026-08-16T12:15:28.709939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.385254Z digest=sha256:28f0d66f297d41a1741cc62d29a11fa84de12dc08a4d32be0fa3d678d1dc2984

Observation 0b433838-7100-49c5-9720-383d7ea431cb · outbound

This paper cites Discovering Language Model Behaviors with Model-Written Evaluations.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Discovering Language Model Behaviors with Model-Written Evaluations

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T12:15:28.390542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:15:28.390542Z digest=sha256:40a85c20da4a563bfc3bba2af6d24fbbd5e0939b5dba85061ea95b629f06a89b

Observation 895217b0-c97e-4883-b5a4-a6d9a35f9dd6 · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 19

Resolution
verified exact
doi, observed 2026-08-16T12:15:28.465603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.396055Z digest=sha256:bc24628259fc0ee7a74a1b22eb095f99c0c64f8a444ba175c4c9d84395aa854e

Observation 7055e07b-53bd-4111-b292-c84552f282c6 · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:15:28.833676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.401386Z digest=sha256:f94f99a179ce76d7bee7c2a3eff670ea39d49fd44997c77272c48b1c2574f9a4

Observation 6111db8b-1807-43d5-9105-aa19b400e5f5 · outbound

This paper cites FPUS23: An Ultrasound Fetus Phantom Dataset with Deep Neural Network Evaluations for Fetus Orientations, Fetal Planes, and Anatomical Features.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability FPUS23: An Ultrasound Fetus Phantom Dataset with Deep Neural Network Evaluations for Fetus Orientations, Fetal Planes, and Anatomical Features

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-16T12:15:28.622059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.406245Z digest=sha256:890a61452fcdb1978f2074583a56b6d40d500899e5af8cbb8ab7c81dea24a515

Observation 65267fd9-6766-40cf-941f-d47787d383ef · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:15:28.818417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.410918Z digest=sha256:7d1a274e3f4aac5fa4526e2a2ede2a3dd9a8df7f9d327ac60b69235e5c37ba67

Observation 9046bc5f-a64d-45b5-a6c1-aa41c35b6b85 · outbound

This paper cites Optimizing Irregular Communication with Neighborhood Collectives and Locality-Aware Parallelism.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Optimizing Irregular Communication with Neighborhood Collectives and Locality-Aware Parallelism

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-16T12:15:28.600797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.415411Z digest=sha256:ee819f701bf2d387a4fb1e24afed28ef302d8dee3a92c3964dc3706d4f6b444f

Observation cd936bbe-a643-44cf-be5b-5116ce1f1024 · outbound

This paper cites The global solution of the minimal surface flow and translating surfaces.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability The global solution of the minimal surface flow and translating surfaces

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-16T12:15:28.577097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.420162Z digest=sha256:c77639630fe34d0abcd7c26fa6196d9c27251ac9b5d7588c939d6f69148b0660

Observation 457fad4e-4b27-4676-a535-046123e9184a · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:15:28.804202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T12:15:28.425038Z digest=sha256:3cc175a6ca7bfb2bb811f2263ed4a714b57b78f0aab71e20631f9391ec90caae

Observation c8a39540-95be-4aea-be97-17c786ab056b · outbound

This paper cites an unresolved cited work.

Governance Challenges in Reinforcement Learning from Human Feedback: Evaluator Rationality and Reinforcement Stability Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T12:15:28.429441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:15:28.429441Z digest=sha256:0af17fe748067f59d543083b45d87f880e896114a1b7443bc62b9f636cf132c0

Pith citing papers

No inbound Pith citation observations are available.