Pith. sign in

Paper Citation Record · LEDGER

Reward-Robust RLHF in LLMs

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2409.15360.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.15360 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:30:36.307631Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T01:37:30.559409Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 213e1a0e-5b42-4517-ad6c-342bee73de6c · inbound

Knowledge Boundary of Large Language Models: A Survey cites this paper.

Knowledge Boundary of Large Language Models: A Survey Reward-Robust RLHF in LLMs

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T14:07:14.401059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:07:14.401059Z digest=sha256:a9afa1da831346fa508f29c7cbfcdd46665d789c329b1f59a46ba653ca19bd66

Observation 127c338a-825f-433d-9939-9d844b6dc90d · inbound

Baichuan4-Finance Technical Report cites this paper.

Baichuan4-Finance Technical Report Reward-Robust RLHF in LLMs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T13:55:19.103768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:55:19.103768Z digest=sha256:3f8b684b98aa028ccf958fb2053cc92e46d536fadaab22fc784f9af67f0c8999

Observation c15e9c54-7036-46d7-8e33-9a41296b7a8d · inbound

Ethics and Persuasion in Reinforcement Learning from Human Feedback: A Procedural Rhetorical Approach cites this paper.

Ethics and Persuasion in Reinforcement Learning from Human Feedback: A Procedural Rhetorical Approach Reward-Robust RLHF in LLMs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:30:36.307631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:30:36.307631Z digest=sha256:3808b88b8ca46ca057a8b3e773d16cf6b4e779fbf39af50bee1225a11ed15910

Observation 1327b6b7-25c7-43f0-9cc9-2622a2a96aa8 · inbound

Bradley-Terry and Multi-Objective Reward Modeling Are Complementary cites this paper.

Bradley-Terry and Multi-Objective Reward Modeling Are Complementary Reward-Robust RLHF in LLMs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T18:51:28.248727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:51:28.248727Z digest=sha256:6693f876254f37af34bc142f006aab8de37abf66925ba8d3ac448f35fe03d13a

Observation b7b578d7-c0c7-49b4-ba62-867ceef01ac0 · inbound

Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges cites this paper.

Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges Reward-Robust RLHF in LLMs

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:00:28.126800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T13:58:53.430492Z digest=sha256:8a76410248e6ac58fe955e1170eb50f2e712460ae5130757e6726154dd8ef92e

Observation a401b329-e1b9-43e1-9961-592a1ba8383b · inbound

Learning to Control Summaries with Score Ranking cites this paper.

Learning to Control Summaries with Score Ranking Reward-Robust RLHF in LLMs

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:41:36.374731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-10T06:41:27.954392Z digest=sha256:6fc81cebe4ee9e751ea4d488b1eebdb3315f7c8d86aeca0ad20d730c2555dfbb

Observation 1945aa31-450c-4866-a9e7-61e18991e1c2 · inbound

Wasserstein Distributionally Robust Regret Optimization for Reinforcement Learning from Human Feedback cites this paper.

Wasserstein Distributionally Robust Regret Optimization for Reinforcement Learning from Human Feedback Reward-Robust RLHF in LLMs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:46:46.441189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-09T21:03:53.045304Z digest=sha256:925bde5cb2cdfb05c5967a77dcf176a7251519e84d80edc7791f546f007a21cf

Observation 8491470d-6bcd-448d-9d64-fda26e931bf3 · inbound

PERSA: Reinforcement Learning for Professor-Style Personalized Feedback with LLMs cites this paper.

PERSA: Reinforcement Learning for Professor-Style Personalized Feedback with LLMs Reward-Robust RLHF in LLMs

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:51:42.415613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-09T19:09:07.557773Z digest=sha256:f4011a3076508ab9919b18b23b5baf4951255be8e3f48da1ecc3888b7743aa29

Observation e6e893e5-515f-43ea-b111-8a29f7588a23 · inbound

Efficient Preference Poisoning Attack on Offline RLHF cites this paper.

Efficient Preference Poisoning Attack on Offline RLHF Reward-Robust RLHF in LLMs

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T00:55:12.741490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-30T23:53:37.665295Z digest=sha256:d3948a6cfaeeee21aa1ecdd51f8af8ba8474ae056b8736e127e5cb79d941f041

Observation 02fd9cc7-723c-42cf-815a-bb74bb9a983a · inbound

Quantifying Empirical Compute-Supervision Tradeoffs in RLVR cites this paper.

Quantifying Empirical Compute-Supervision Tradeoffs in RLVR Reward-Robust RLHF in LLMs

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T11:54:38.520665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-30T11:46:34.688973Z digest=sha256:b4e81cad2ba2622c9c7917b118c37462af7e4fecee8a997642e4b5cd75848be0

Observation 7b1698c6-85d1-4d09-996a-25ae3ab407e6 · inbound

A Unifying Lens on Reward Uncertainty in RLHF cites this paper.

A Unifying Lens on Reward Uncertainty in RLHF Reward-Robust RLHF in LLMs

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T00:27:29.818495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-27T17:11:46.549150Z digest=sha256:a9af36953a81a71de6f50d35f74d50805d9023575318d8b296ff6a15da6888fe

Observation fdc39863-b19b-451a-9a85-2277f186b268 · inbound

The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning cites this paper.

The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning Reward-Robust RLHF in LLMs

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:37:30.062012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-27T17:07:06.227106Z digest=sha256:31636e2006885e7bd29ad7258e9ebb84e2ec7cf73f73eedf2e04d6a0eb96ec74

Observation c20bfd58-f270-4d15-b8ef-ee2bed78555f · inbound

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization cites this paper.

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization Reward-Robust RLHF in LLMs

Reference 233

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:37:30.560683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-27T16:26:34.918099Z digest=sha256:46822794aa8d1749b1eb29fb0ad878189914c5b7082701dda0a4654792e7580d

Observation be211758-972e-4e32-bb4f-70a2031c0cdf · inbound

Multimodal Reward Hacking in Reinforcement Learning cites this paper.

Multimodal Reward Hacking in Reinforcement Learning Reward-Robust RLHF in LLMs

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-13T02:39:02.891861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T02:39:02.891861Z digest=sha256:1a86db488f1d62cfcbe6628ea72e477450751a15c723abc4bc43d70c5a6c9d76

Observation bd5acc56-f3af-44d1-b7b0-0daffa70db87 · inbound

SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text cites this paper.

SCOPE and SCION: A Benchmark and an Auditable Reference Pipeline for Schema Induction and Fusion from Text Reward-Robust RLHF in LLMs

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T13:36:57.114439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T13:36:57.114439Z digest=sha256:c133ccb9888fd68b0387ec021496d69e701a1395eb171f4e4cab9b5c39dffe81

Observation 54b01754-a1cc-4854-a647-b298e0e23486 · inbound

Robust General Utility for Reinforcement Learning cites this paper.

Robust General Utility for Reinforcement Learning Reward-Robust RLHF in LLMs

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T14:55:08.347331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:55:08.347331Z digest=sha256:50db39de59143865ecaad86519ae5183bbac11eb4636e9a25013981571140d9d