Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-08T21:20:07.569587Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 2 inbound Pith citation observations for arXiv:2607.05904.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-08T21:20:07.569587Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T05:30:04.547511Z
A source-named dated measurement, never combined with another source.
Source: cited_works
17 of 17 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 506e4a48-45dd-4232-bdb2-632c0437beae · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges Constitutional AI: Harmlessness from AI Feedback
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a2a38675-5ead-4f4a-89e7-1e1b34693671 · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e426a20c-f573-4c4b-a2db-4c6fccabcff7 · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges Reward Model Ensembles Help Mitigate Overoptimization
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6539054c-9dac-4621-b592-f3be852adeb6 · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges Scaling Laws for Reward Model Overoptimization
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5f5cc307-ba46-4aab-bcad-d2edb1acc691 · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges On scalable oversight with weak LLMs judging strong LLMs
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3aa99a92-a834-4240-aaea-2d299ae90004 · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges When Does Verification Pay Off? A Closer Look at LLMs as Solution Verifiers
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 69d997df-802e-42d1-aafe-6cd36dd11edf · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges Spontaneous Reward Hacking in Iterative Self-Refinement
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c8f2e35b-8718-4d47-a563-965679ca8cff · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges Scaling Laws for Reward Model Overoptimization in Direct Alignment Algorithms
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation bb7db7e6-05e1-4192-86a6-93676208df0a · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges Towards Understanding Sycophancy in Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 74981873-42cf-49ce-8db6-a0512b782060 · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges RLSR: Reinforcement Learning from Self Reward
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation fad3a143-a4bd-4dc4-9ba6-2eac61ce336e · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges A Long Way to Go: Investigating Length Correlations in RLHF
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8ecccce0-e3b5-4121-8e76-605d931b0bb1 · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges Language Models Learn to Mislead Humans via RLHF
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3f3cf442-54df-48bf-9996-6e8ce059e79d · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges Self-Rewarding Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 673ffde6-90b0-4138-9369-886aaa53056e · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges Generative Verifiers: Reward Modeling as Next-Token Prediction
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 23382bfc-ab7b-4f1b-95f5-22557924239b · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9f584040-5632-4e7d-8caa-742fc0662be9 · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges My answer:
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation dd5d18f7-6d21-4e60-851b-7e93d26be757 · outbound
More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6ae011a9-146e-4fe1-beee-c62444a77711 · inbound
When the Reward Suite Is Leaky: A Preregistered Causal Contrast of Natural Verifier False Positives in RLVR More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c00c37c5-f72f-4540-aa54-4c93eb9684bd · inbound
LLM-as-a-Judge Scores Are Unreliable Optimization Signals in Closed-Loop Table Recognition More Convincing, Not More Correct: Self-Play Reward Hacking of Reference-Free LLM Judges
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.