Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:54:06.092545Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 7 inbound Pith citation observations for arXiv:2504.14439.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:54:06.092545Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T23:03:37.442512Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T20:57:23.584256Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ecf0a54b-fd87-4174-92b9-a35be6c03888 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d399bfb9-26bd-47a1-bffc-0a0f0140d786 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling GPT-4 Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9105295-c371-4a3f-b679-62062fe2e96d · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Flambe: Structural complexity and representation learning of low rank mdps
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 73b1d9de-15c3-4d27-b973-0b783d1c941d · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Constitutional AI: Harmlessness from AI Feedback
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a072a21-c696-4576-b5e2-2eba6fa53f40 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Fine-tuning language models to find agreement among humans with diverse preferences
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4b684668-6050-433d-9d0f-39696fea3082 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Initializing services in interactive ml systems for diverse users
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4a0a655c-2dc4-414f-9518-5b358487efb0 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Offline Multi-task Transfer RL with Representational Penalization
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fcf749b-26c5-4c0f-9315-7ee6921fc679 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Rank analysis of incomplete block designs: I
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2546c224-e0dd-4536-82a1-34fac97c789c · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling PERSONA: A Reproducible Testbed for Pluralistic Alignment
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 272ed6ce-6d50-4c36-84ba-5d039cab456e · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling PAL: Pluralistic Alignment Framework for Learning from Heterogeneous Preferences
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29c1425a-709c-4f17-902c-3461825f61b2 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling PAD: Personalized Alignment of LLMs at Decoding-Time
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07f6830c-107b-4507-afb7-e34a156802b7 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Deep reinforcement learning from human preferences
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a2d977d-04bc-4635-a84d-727dbd7d72ff · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Can LLM be a Personalized Judge?
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4158e063-c351-437c-8ff1-ba7c81a359db · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Towards Measuring the Representation of Subjective Global Opinions in Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c783997-b118-4c07-92a1-28f4c42c5747 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Value Augmented Sampling for Language Model Alignment and Personalization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12e45069-22f4-4e29-87e3-ae74d91a6009 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling LoRA: Low-Rank Adaptation of Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 177894a9-0402-4fd6-b48d-aafc74e811a1 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b329dfcb-eaf8-41be-be61-49db1bf9d1e3 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Evaluating and inducing personality in pre-trained language models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 335a34c4-273e-4efb-b9e9-0006750a9b4b · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Adam: A Method for Stochastic Optimization
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a4ad103-f726-4f08-ad07-a6cca18d0967 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling The benefits, risks and bounds of personalizing the alignment of large language models to individuals
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 93b7d558-ade3-418f-9336-d19af72f863d · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68ab15d1-13d9-4bc3-ac34-0951fa252e90 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Matrix factorization techniques for recommender systems
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8377d5c3-a141-483a-9804-041723d24820 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Evaluating Cultural Adaptability of a Large Language Model via Simulation of Synthetic Personas
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6a0b2b9-bfde-4af9-983b-d7eb48007b6e · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Test-Time Alignment via Hypothesis Reweighting
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1aa47a59-427c-4e67-ad8b-f266a9a67c87 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Personalized Language Modeling from Personalized Human Feedback
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 184c3746-fa6f-4ca7-b058-57e033fe3adc · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d0161a1-d6b8-4999-830b-4e912285901c · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Individualized rank aggregation using nuclear norm regularization
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation abb3ef21-f5da-4df4-96fa-eb2f2090f8e1 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Virtual Personas for Language Models via an Anthology of Backstories
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3549ba9-2fb4-4104-b393-5bf2e3269153 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Training language models to follow instructions with human feedback
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 425c773c-24a3-440b-80cf-e1d5f7f19165 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Llm evaluators recognize and favor their own generations
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b345a217-bffd-45a1-8d34-4d25666272f4 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Preference completion: Large-scale collaborative ranking from pairwise comparisons
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 77042ddd-8aaf-49e0-96ce-7b22889c7c44 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c395793e-182f-4d4e-a0d2-f44fcb5d711c · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Direct preference optimization: Your language model is secretly a reward model
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5133b998-d4c5-4c5a-9f66-d281e024b579 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Rewarded soups: towards pareto-optimal alignment by interpolating weights fine-tuned on diverse rewards
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48eddf9c-1cc5-43bc-aebc-155e6341d9c4 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Personality traits in large language models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation bdd2785c-06b1-401a-9e49-ddc3613e496e · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Aligning Language Models with Demonstrated Feedback
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a18016d9-115a-49ed-92ec-fac52e27ea8b · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Decoding-Time Language Model Alignment with Multiple Objectives
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2ee5bd2-3e0a-464b-bf55-98169421c3b9 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling FSPO: Few-Shot Optimization of Synthetic Preferences Personalizes to Real Users
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9deda73d-a1ad-4809-9739-c4f638485e67 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling A Roadmap to Pluralistic Alignment
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55330826-cfa2-4c37-b4a1-f01887116134 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Learning to summarize with human feedback
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9df6f9b7-735c-41da-9f6d-223ad81b8861 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling LLaMA: Open and Efficient Foundation Language Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87d360a2-0c1b-4601-8fdd-e8a7c92f2419 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Interpretable Preferences via Multi-Objective Reward Modeling and Mixture-of-Experts
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ec9883b-f703-4de7-8738-d1dab63c1356 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Personalized Large Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8333ac3-9a37-4bda-9dea-f977ff39a6ff · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Fine-grained human feedback gives better rewards for language model training
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3ef44be-5a1f-4b90-b46c-6a02a921506b · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Group Preference Optimization: Few-Shot Alignment of Large Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e0f880f-7f83-43c9-b6a0-7e51c2b5d8d5 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Beyond One-Preference-Fits-All Alignment: Multi-Objective Direct Preference Optimization
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a92bf4e-2c36-4cea-83a9-2b220a95fcdc · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Personality Alignment of Large Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ac1b61b-3fde-401d-966f-f361b7bee87b · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling Fine-Tuning Language Models from Human Preferences
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e095462-30f9-4058-aa0f-73cd05767539 · outbound
LoRe: Personalizing LLMs via Low-Rank Reward Modeling PersonalLLM: Tailoring LLMs to Individual Preferences
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a903f797-26e0-48dc-b27f-1e35298c32d9 · inbound
A Personalized Conversational Benchmark: Towards Simulating Personalized Conversations LoRe: Personalizing LLMs via Low-Rank Reward Modeling
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea4d2761-f2f0-4393-b499-bf9afbc11de8 · inbound
Preference Learning for AI Alignment: a Causal Perspective LoRe: Personalizing LLMs via Low-Rank Reward Modeling
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1c4b4c45-1162-4a68-aaa7-3627c71f13cc · inbound
POPI: Personalizing LLMs via Optimized Natural Language Preference Inference LoRe: Personalizing LLMs via Low-Rank Reward Modeling
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation aa7e733a-bcb4-4678-b3b2-871a213e889e · inbound
LLM Evaluation as Tensor Completion: Low Rank Structure and Semiparametric Efficiency LoRe: Personalizing LLMs via Low-Rank Reward Modeling
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation c59bf9d7-60d2-4694-bb6c-789bb9dc1565 · inbound
Beyond Isolated Behaviors: Hierarchical User Modeling for LLM Personalization LoRe: Personalizing LLMs via Low-Rank Reward Modeling
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e83153ec-6383-49ae-9ca9-d9a4c4ab2921 · inbound
PAFO: Pareto Fairness Optimization for Personalized Reward Modeling LoRe: Personalizing LLMs via Low-Rank Reward Modeling
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 78095518-2691-4c04-936b-dde539db32b6 · inbound
Cautious Context Steering for Language Model Personalization LoRe: Personalizing LLMs via Low-Rank Reward Modeling
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.