Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2408.10075.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:39:52.380908Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T20:58:57.620585Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation cb318282-0a32-42fa-b0e6-2834b0a3daba · inbound
Test-Time Alignment via Hypothesis Reweighting Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 46fd3c2e-fbc4-4ea4-ae96-ca85e99b6fc1 · inbound
Extended Inductive Reasoning for Personalized Preference Inference from Behavioral Signals Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d1d615b-98dc-469c-a8cf-fc480c450fee · inbound
Configurable Preference Tuning with Rubric-Guided Synthetic Data Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9ac34ef-cfd6-43e6-b9fc-2a83d87898d4 · inbound
Affordances of Sketched Notations for Multimodal UI Design and Development Tools Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 170584b6-cd83-4933-b97a-d3ddcbacea7a · inbound
SharedRep-RLHF: A Shared Representation Approach to RLHF with Diverse Preferences Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32433d29-692a-4ba9-8ed1-d2402c47067e · inbound
CURP: Codebook-based Continuous User Representation for Personalized Generation with LLMs Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Reference 2003
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 523a3b92-9d22-4554-bfec-28da7f1a675d · inbound
Efficient Personalization of Generative User Interfaces Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 22b13f78-f74c-4dae-859f-b921d38188f6 · inbound
Federated Variational Preference Alignment with Gumbel-Softmax Prior for Personalized User Preferences Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fea0eeed-0edb-4c58-b0ae-098c20161292 · inbound
Beyond Isolated Behaviors: Hierarchical User Modeling for LLM Personalization Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ac38de3a-4910-453d-87e6-e49889b16602 · inbound
Using Cognitive Models to Improve Language Model Simulation of Human Persuasion Games Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6541a1d8-99be-4956-8599-770a848228d5 · inbound
Personalizing Large Language Model Agents with Small Policy Models Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.