Pith. sign in

Paper Citation Record · LEDGER

Reward-Robust RLHF in LLMs

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2409.15360.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.15360 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:30:36.307631Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T01:37:30.559409Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 213e1a0e-5b42-4517-ad6c-342bee73de6c · inbound

Knowledge Boundary of Large Language Models: A Survey cites this paper.

Knowledge Boundary of Large Language Models: A Survey Reward-Robust RLHF in LLMs

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T14:07:14.401059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:07:14.401059Z digest=sha256:f019e940ed79b2b607905a3f3c7cadfa28bc40edd702c10d59c87c63a40bc855

Observation 127c338a-825f-433d-9939-9d844b6dc90d · inbound

Baichuan4-Finance Technical Report cites this paper.

Baichuan4-Finance Technical Report Reward-Robust RLHF in LLMs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T13:55:19.103768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:55:19.103768Z digest=sha256:d6e25289bde1c6be5bae321922116566db87baa6614e68f59cc37e67c639bc39

Observation c15e9c54-7036-46d7-8e33-9a41296b7a8d · inbound

Ethics and Persuasion in Reinforcement Learning from Human Feedback: A Procedural Rhetorical Approach cites this paper.

Ethics and Persuasion in Reinforcement Learning from Human Feedback: A Procedural Rhetorical Approach Reward-Robust RLHF in LLMs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:30:36.307631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:30:36.307631Z digest=sha256:ac101fe9a16f94cd65164500878fbd581ea01f0faf246a060cae7290c71ca4dd

Observation 1327b6b7-25c7-43f0-9cc9-2622a2a96aa8 · inbound

Bradley-Terry and Multi-Objective Reward Modeling Are Complementary cites this paper.

Bradley-Terry and Multi-Objective Reward Modeling Are Complementary Reward-Robust RLHF in LLMs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:51:28.248727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:51:28.248727Z digest=sha256:a306174fae7dafaf9ce09e39a669dcd67b74f401f7ff8a5818ffa00b8ae80714

Observation b7b578d7-c0c7-49b4-ba62-867ceef01ac0 · inbound

Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges cites this paper.

Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges Reward-Robust RLHF in LLMs

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:00:28.126800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T13:58:53.430492Z digest=sha256:adba99651a6f84adca57afc50f28be9a03afe74214419d4202662a1d4fb35a16

Observation a401b329-e1b9-43e1-9961-592a1ba8383b · inbound

Learning to Control Summaries with Score Ranking cites this paper.

Learning to Control Summaries with Score Ranking Reward-Robust RLHF in LLMs

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:41:36.374731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-10T06:41:27.954392Z digest=sha256:de0af9480116f3aa1554c73aba50a3553a3837ea42a9cfcba442bda3311c5c8d

Observation 1945aa31-450c-4866-a9e7-61e18991e1c2 · inbound

Wasserstein Distributionally Robust Regret Optimization for Reinforcement Learning from Human Feedback cites this paper.

Wasserstein Distributionally Robust Regret Optimization for Reinforcement Learning from Human Feedback Reward-Robust RLHF in LLMs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:46:46.441189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-09T21:03:53.045304Z digest=sha256:6e50c9ab909ae99b1a50928c328b61c308dda8d286a116f9723a5fa426eb2e82

Observation 8491470d-6bcd-448d-9d64-fda26e931bf3 · inbound

PERSA: Reinforcement Learning for Professor-Style Personalized Feedback with LLMs cites this paper.

PERSA: Reinforcement Learning for Professor-Style Personalized Feedback with LLMs Reward-Robust RLHF in LLMs

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:51:42.415613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-09T19:09:07.557773Z digest=sha256:c2577836deccb5a0f1cfed76360502a4ccb0f8007a464b070de3f8251249d6f5

Observation e6e893e5-515f-43ea-b111-8a29f7588a23 · inbound

Efficient Preference Poisoning Attack on Offline RLHF cites this paper.

Efficient Preference Poisoning Attack on Offline RLHF Reward-Robust RLHF in LLMs

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T00:55:12.741490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T23:53:37.665295Z digest=sha256:a4a2895da3557d98ee0647a866302ac478761dcdc8fc575291cc54983c350620

Observation 02fd9cc7-723c-42cf-815a-bb74bb9a983a · inbound

Quantifying Empirical Compute-Supervision Tradeoffs in RLVR cites this paper.

Quantifying Empirical Compute-Supervision Tradeoffs in RLVR Reward-Robust RLHF in LLMs

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T11:54:38.520665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T11:46:34.688973Z digest=sha256:a1588aca57c47358e2bff73efb06c2b57c546f2e8156b40175902c3ee94a5fa7

Observation 7b1698c6-85d1-4d09-996a-25ae3ab407e6 · inbound

A Unifying Lens on Reward Uncertainty in RLHF cites this paper.

A Unifying Lens on Reward Uncertainty in RLHF Reward-Robust RLHF in LLMs

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T00:27:29.818495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T17:11:46.549150Z digest=sha256:0fb8c92f248985989228988dc8a56c46343b4d19fd0e084e3cd7c9790c1946d5

Observation fdc39863-b19b-451a-9a85-2277f186b268 · inbound

The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning cites this paper.

The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning Reward-Robust RLHF in LLMs

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:37:30.062012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T17:07:06.227106Z digest=sha256:794cf202fe2bf80218a13487852a7687b9639565f729a49c697c4ef71ecad554

Observation c20bfd58-f270-4d15-b8ef-ee2bed78555f · inbound

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization cites this paper.

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization Reward-Robust RLHF in LLMs

Reference 233

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:37:30.560683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T16:26:34.918099Z digest=sha256:78538b010d51b12ea807292b313de0d6e8e7bfc199f1a0c82aed6b1674ab49f7

Observation be211758-972e-4e32-bb4f-70a2031c0cdf · inbound

Multimodal Reward Hacking in Reinforcement Learning cites this paper.

Multimodal Reward Hacking in Reinforcement Learning Reward-Robust RLHF in LLMs

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-13T02:39:02.891861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T02:39:02.891861Z digest=sha256:71c18ae6e23f5e86437831c3238ed7a58d90f7c5d93a277e6b1b89580101214b

Observation bd5acc56-f3af-44d1-b7b0-0daffa70db87 · inbound

SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text cites this paper.

SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text Reward-Robust RLHF in LLMs

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T13:36:57.114439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T13:36:57.114439Z digest=sha256:d462f9635ee4fde37ba8cf666f1adc155c4717464dda3a148bc0c0e53382d59b

Observation 54b01754-a1cc-4854-a647-b298e0e23486 · inbound

Robust General Utility for Reinforcement Learning cites this paper.

Robust General Utility for Reinforcement Learning Reward-Robust RLHF in LLMs

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T14:55:08.347331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:55:08.347331Z digest=sha256:fa2a8671628a303f57b69e38728818e64fef979493a25d8e6dd1bb8da66855c3