Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:31:05.857015Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 1 inbound Pith citation observation for arXiv:2505.20336.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:31:05.857015Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-10T20:33:46.647842Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
15 of 15 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 06d0a058-c3cb-4990-abfd-c649518d201e · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82af8d40-cac6-48eb-9022-41cfdd1e25c5 · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification UltraFeedback: Boosting Language Models with Scaled AI Feedback
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70ae6733-b1b2-4dfb-88e2-ee7e74a226c5 · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification Direct Language Model Alignment from Online AI Feedback
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10a44876-4afa-4080-af6c-b2b9eea30c26 · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification Aligning to Thousands of Preferences via System Message Generalization
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62f35081-5fa3-4a89-8e4b-d5af209df259 · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification Aligning Crowd Feedback via Distributional Preference Reward Modeling
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 44297587-a9cb-4257-854d-c403c37fff08 · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification Training language models to follow instructions with human feedback
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcc5c1fd-e4c7-47db-96c4-153fe6678fa9 · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f5dbaf0-6f40-460e-90b1-29500f891c01 · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification Ignore This Title and HackAPrompt: Exposing Systemic Vulnerabilities of LLMs through a Global Scale Prompt Hacking Competition
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b969249-ef59-424f-acf7-a4c47ca8d4b7 · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b27f801b-3457-4fb5-8e2f-571b173eed9a · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification In the Table 5 in Ap- pendix B, <preference n > represents a preference intensity of n (1 ≤ n < nmax), where a larger n indicates a higher intensity
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6ce1a30f-7a2f-4ed8-8f29-b91fe03a5eff · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification Learning to summarize from human feedback
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7323d348-b0fc-4847-9203-2a850d839a04 · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification Interpretable Preferences via Multi-Objective Reward Modeling and Mixture-of-Experts
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 310957cd-3dfd-4981-9928-d4c190692c08 · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification RLHF from Heterogeneous Feedback via Personalization and Preference Aggregation
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ee2d541-a903-4e3b-a6b8-2481a81d7726 · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification Rewarded soups: towards Pareto-optimal alignment by interpolating weights fine-tuned on diverse rewards
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 69706b5c-63c6-4108-9952-3e4aa2897442 · outbound
MOSLIM:Align with diverse preferences in prompts through reward classification Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2de5b500-11de-4a72-8e35-5b040521bc9c · inbound
One Model for All: Multi-Objective Controllable Language Models MOSLIM:Align with diverse preferences in prompts through reward classification
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.