Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2402.18571.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T11:20:48.489366Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T07:59:50.977785Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation b5c40558-960e-416e-b402-753a034c8944 · inbound
Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c8814b4-ac09-42a0-8df4-2e87ae403b07 · inbound
OpenReview Should be Protected and Leveraged as a Community Asset for Research in the Era of Large Language Models Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 151
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7c8015e-3c13-4bf1-bfbc-7df13e0bfdea · inbound
CALMA: A Process for Deriving Context-aligned Axes for Language Model Alignment Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 455ce814-39ff-4b23-ac00-f1f8cb8a6233 · inbound
Rubrics as Rewards: Reinforcement Learning Beyond Verifiable Domains Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6f690dbf-7c73-41e2-8399-c55da0741e86 · inbound
Generating Place-Based Compromises Between Two Points of View Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b3524346-517f-4318-8abe-cf6999948d15 · inbound
Response Time Enhances Alignment with Heterogeneous Preferences Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 164
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e1eb872b-f9b3-4eca-b37c-7cdbd9eec21d · inbound
CLIPer: Tailoring Diverse User Preference via Classifier-Guided Inference-Time Personalization Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f5a28b24-336f-4925-9cc1-7de381f28271 · inbound
Explaining and Breaking the Safety-Helpfulness Ceiling via Preference Dimensional Expansion Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7675384e-1375-4dae-9231-c096f722abec · inbound
Explaining and Breaking the Safety-Helpfulness Ceiling via Preference Dimensional Expansion Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1de03b3c-9dc8-4e29-bbd7-4c767030b6f2 · inbound
MOCHA: Multi-Objective Chebyshev Annealing for Agent Skill Optimization Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bb1526a0-e8e7-49a9-8bfd-b38fd5a1973e · inbound
Spectral Souping: A Unified Framework for Online Preference Alignment Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dfaca053-4c67-49e6-b91a-07cf2c2eab44 · inbound
Multi-Turn On-Policy Distillation with Prefix Replay Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 148
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 364a638b-056d-4e49-bf1e-b4ceb9c665a2 · inbound
Multi-Turn On-Policy Distillation with Prefix Replay Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards
Reference 149
Source-reported events for the cited work
Unavailable: canonical work link unavailable.