Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:31:45.734231Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 0 inbound Pith citation observations for arXiv:2501.06248.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:31:45.734231Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
33 of 33 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 63b0bb0d-2444-4081-ac61-41292b01466b · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69553556-f3f6-4317-9aeb-0f9ea8d6161c · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models On the Opportunities and Risks of Foundation Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 056526bc-d7af-4d73-a5af-30a5c2f08331 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Language Models are Few-Shot Learners
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3b3b00e-561a-4824-ad5b-31f4003490f7 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 817eda52-d0f5-431c-bd61-25f2726e6f93 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3abef2ae-4105-4fe2-bd5a-372bbf6d1b1c · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Helping or Herding? Reward Model Ensembles Mitigate but do not Eliminate Reward Hacking
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ef1db20-cc05-4cec-9123-ba7a076ec3f8 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Attention Flows are Shapley Value Explanations
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d43c383f-e8ab-48e3-9731-411802b1b3a6 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Axioms for AI Alignment from Human Feedback
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c32782a-3e91-42a1-9c00-e763f65a879d · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Gemma: Open Models Based on Gemini Research and Technology
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12272422-9d9d-4b21-a3ac-c2ff4ce161fb · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Steering Language Models with Game-Theoretic Solvers
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24db5438-2529-4b7c-8f0c-ccc38c6f2f56 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Improving alignment of dialogue agents via targeted human judgements
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6080f8c3-ea81-4161-ad10-30f11ebf0a5e · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models The Consensus Game: Language Model Generation via Equilibrium Search
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3d976a8-3a21-4061-9a45-7ad1402a7f05 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 23e0c1d8-1691-480e-baa4-8a0aa3aa6c1a · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Confronting Reward Model Overoptimization with Constrained RLHF
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cfdc89b-b94b-4f18-8c90-3004031e06b8 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Nash Learning from Human Feedback
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b4998d3-5525-4a81-abfd-67e7a4ba4e99 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Learning Social Welfare Functions
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 847d409e-a360-4a39-83a3-f8fdf8999ddb · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 771fc87a-285d-4986-bfda-c994e177d059 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b1bd732e-37d6-44c9-aa74-bade1ffe93d1 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ddb1ab5c-24e8-42e3-bc45-a5ee8f733e34 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 65e807e5-2b58-4273-b5a1-8043d6d25af1 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Proximal Policy Optimization Algorithms
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 571b9802-8a55-474e-81c0-c2e5a4ec0f20 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models A Long Way to Go: Investigating Length Correlations in RLHF
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b86165bb-7e5f-458a-a718-31a4d8330e34 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e699ef17-cf16-4f60-a1e9-a2a258580c43 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models A Minimaximalist Approach to Reinforcement Learning from Human Feedback
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aab78833-519c-4606-8f87-c64c2f219b4b · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Evaluating and Mitigating Discrimination in Language Model Decisions
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e46d02ab-af9a-4a41-91e5-acf244c16eeb · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 97accfb3-d797-48ae-9bd5-3cb04c4287d0 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Foundational Autoraters: Taming Large Language Models for Better Automatic Evaluation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a80e9000-c40a-4121-964b-61f1b9dc0c67 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Transforming and Combining Rewards for Aligning Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a26d0694-f2fd-4af0-b501-231962d711fe · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Ethical and social risks of harm from Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c7c0d23-7c6c-47ca-9aef-7828554af0ff · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 146e3566-1b63-4247-bd29-bb6993028f04 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models Fine-Grained Human Feedback Gives Better Rewards for Language Model Training
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a496f50-7965-4c1a-b2dc-cddd7aa76bc5 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models online" 'onlinestring :=
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b515aae-a9cf-4cbc-919a-138da253c004 · outbound
Utility-inspired Reward Transformations Improve Reinforcement Learning Training of Language Models write newline
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.