Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T20:51:53.299361Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2411.09341.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T20:51:53.299361Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
35 of 35 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f6544a8d-a5d6-45ec-9049-433565394c03 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment , " * write output.state after.block = add.period write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec330355-af13-4987-84ff-9aec71cb526b · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e711594-8e85-4c93-9b0d-0a169ca5e309 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment GPT-4 Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07a35865-d9b9-46b2-a8eb-14a626b7c5be · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af6357ba-414a-47e1-b329-5a5bf6b44c93 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment A.; and Terry, M
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 913daf76-fb90-4b3e-b938-856f51740ebc · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment J.; and van der Schaar, M
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 79fa944b-b6a5-4a26-a5ec-b35439a72899 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Reward Model Ensembles Help Mitigate Overoptimization
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0a04680-6b45-42a7-9a5b-3c3d099f4132 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dd6691b-88d4-4ef4-a09f-c38caea325b8 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Dissecting Recall of Factual Associations in Auto-Regressive Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 180e9686-c04d-49b6-9e74-8d0dc51b390b · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6a1a0f39-1a00-4596-90d0-c3ecab994e9b · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment The Curious Case of Neural Text Degeneration
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d11b751f-1c02-46ca-9c8e-d1027f6df6b0 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Preference Transformer: Modeling Human Preferences using Transformers for RL
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4390c02-6ff5-4328-ba88-08fc5c607b18 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Playing Atari with Deep Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3fd12f4-8589-4611-b78f-e21d994e7d5c · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment WebGPT: Browser-assisted question-answering with human feedback
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f3beb00-962d-4a3c-95e5-dc13be36a4b0 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Y.; Russell, S.; et al
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a410275f-acb1-49cb-8dd7-7e14af80026c · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ede472ff-1e61-4887-b132-1679e2df341b · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92eb5f4e-8d06-4743-acbe-ff91a888c7b5 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment D.; Ermon, S.; and Finn, C
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59597ecd-d41a-4535-a1cf-a4ef9e0c76e0 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d466063-c013-4035-a85b-dba08fc55272 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a4512eda-ed6d-4102-8cf3-c51b8e3cd1fa · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Proximal Policy Optimization Algorithms
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88c54870-04ac-4b88-88ed-0e746332ae9e · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Large Language Model Alignment: A Survey
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64667873-9296-4c99-8cca-8b6df2649084 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d98f7b19-1569-4d91-a1bd-3423859b9ecb · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e5b2aec-a4b9-4f0f-a803-cab55c4a8a75 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b34079ab-483f-49e5-9763-b1af369a2c45 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Inverse-RLignment: Large Language Model Alignment from Demonstrations through Inverse Reinforcement Learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17c4dcee-5634-4571-a3c4-435e1c433264 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment T.; and Pal, C
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1f864896-899f-46a8-849d-b6da04f8975d · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ba8518f-d633-4edd-ab5b-951a2990b1d9 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment N.; Kaiser, .; and Polosukhin, I
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23ffa5b9-7eaa-41c1-83b8-fbf772db50e8 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Ethical and social risks of harm from Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 853c1f88-dd45-4e82-9661-9e8fb6cff1a9 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Baichuan 2: Open Large-scale Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dbe5ef3-bce9-423e-8da1-daa8ebc825c6 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment RRHF: Rank Responses to Align Language Models with Human Feedback without tears
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64db0c9c-9e6c-4e0c-b0ab-cb8c096fa715 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Overcoming Reward Overoptimization via Adversarial Policy Optimization with Lightweight Uncertainty Estimation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59bd2236-d0c8-4a6b-bd09-72a228229ad7 · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment DialoGPT: Large-Scale Generative Pre-training for Conversational Response Generation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8e0967c-79b7-4f8f-bf22-c139f106c22d · outbound
Approximated Variational Bayesian Inverse Reinforcement Learning for Large Language Model Alignment Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
No inbound Pith citation observations are available.