Pith. sign in

Paper Citation Record · LEDGER

Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2402.18571.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.18571 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T11:20:48.489366Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T07:59:50.977785Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b5c40558-960e-416e-b402-753a034c8944 · inbound

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants cites this paper.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:48.489366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:48.489366Z digest=sha256:a9aa692adbf0f2ed6256097e573cf0bba715cd67d12717a3fd254575c36cda57

Observation 2c8814b4-ac09-42a0-8df4-2e87ae403b07 · inbound

OpenReview Should be Protected and Leveraged as a Community Asset for Research in the Era of Large Language Models cites this paper.

OpenReview Should be Protected and Leveraged as a Community Asset for Research in the Era of Large Language Models Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 151

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:47.967756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:31:47.967756Z digest=sha256:a90beed6c40782a07b565b5762a58467081bf68e784fcd9b92f538803368e653

Observation e7c8015e-3c13-4bf1-bfbc-7df13e0bfdea · inbound

CALMA: A Process for Deriving Context-aligned Axes for Language Model Alignment cites this paper.

CALMA: A Process for Deriving Context-aligned Axes for Language Model Alignment Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T18:10:18.299705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:10:18.299705Z digest=sha256:432b9cf7156e41804c7f5c3b5807113680d4fd92e767aed1a984bdee3668f3c5

Observation 455ce814-39ff-4b23-ac00-f1f8cb8a6233 · inbound

Rubrics as Rewards: Reinforcement Learning Beyond Verifiable Domains cites this paper.

Rubrics as Rewards: Reinforcement Learning Beyond Verifiable Domains Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:07:56.859644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T06:07:56.678339Z digest=sha256:7493f471dbb52c651536272ce7c80b4f5d6c7084cedb09ee4e1b676845d2c8d9

Observation 6f690dbf-7c73-41e2-8399-c55da0741e86 · inbound

Generating Place-Based Compromises Between Two Points of View cites this paper.

Generating Place-Based Compromises Between Two Points of View Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:01:12.114202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T03:36:31.695964Z digest=sha256:8ab2cae8448f3f44efa972b011eeabd41159a059a7e4ee9d4eaa66136fbc5371

Observation b3524346-517f-4318-8abe-cf6999948d15 · inbound

Response Time Enhances Alignment with Heterogeneous Preferences cites this paper.

Response Time Enhances Alignment with Heterogeneous Preferences Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 164

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:45:59.808702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T01:04:26.288913Z digest=sha256:b4b6809327a72cff6588ffc3e09b18156ef545532f8cc2c85defd34e486dfb31

Observation e1eb872b-f9b3-4eca-b37c-7cdbd9eec21d · inbound

CLIPer: Tailoring Diverse User Preference via Classifier-Guided Inference-Time Personalization cites this paper.

CLIPer: Tailoring Diverse User Preference via Classifier-Guided Inference-Time Personalization Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T03:50:54.299657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T02:16:28.593349Z digest=sha256:5f7072249160146d55eca57e42d6d8ec85208dfdee3ab9d083c069fe93fe43a0

Observation f5a28b24-336f-4925-9cc1-7de381f28271 · inbound

Explaining and Breaking the Safety-Helpfulness Ceiling via Preference Dimensional Expansion cites this paper.

Explaining and Breaking the Safety-Helpfulness Ceiling via Preference Dimensional Expansion Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:57:06.526224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T01:03:10.263663Z digest=sha256:7fbc5ec1f42209acfb82c6a2992a8d9dcc645c5fbccd9b59db799ffc57ebd985

Observation 7675384e-1375-4dae-9231-c096f722abec · inbound

Explaining and Breaking the Safety-Helpfulness Ceiling via Preference Dimensional Expansion cites this paper.

Explaining and Breaking the Safety-Helpfulness Ceiling via Preference Dimensional Expansion Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:12:58.887870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T21:12:06.989077Z digest=sha256:11acca2488fc6082e0b7e0c029931d9df465353b5df40950732e617b231296e0

Observation 1de03b3c-9dc8-4e29-bbd7-4c767030b6f2 · inbound

MOCHA: Multi-Objective Chebyshev Annealing for Agent Skill Optimization cites this paper.

MOCHA: Multi-Objective Chebyshev Annealing for Agent Skill Optimization Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:13:05.313131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T06:09:56.684622Z digest=sha256:208f68831cd48c6c201445cf260a532c912f4d459aa2ce7519f60fded31a0b4d

Observation bb1526a0-e8e7-49a9-8bfd-b38fd5a1973e · inbound

Spectral Souping: A Unified Framework for Online Preference Alignment cites this paper.

Spectral Souping: A Unified Framework for Online Preference Alignment Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:59:50.979793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T07:54:56.356555Z digest=sha256:8fbd12856ec0549c6381327bd9e464dd7a665f11ba852910d9d941d3bba6c992

Observation dfaca053-4c67-49e6-b91a-07cf2c2eab44 · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 148

Resolution
unresolved
no resolver link, observed 2026-07-11T13:53:36.775836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T13:53:36.775836Z digest=sha256:52b50f6994672863abcf347254090a08d117a18affe83d1bda387cb4f0d95fcc

Observation 364a638b-056d-4e49-bf1e-b4ceb9c665a2 · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 149

Resolution
unresolved
no resolver link, observed 2026-08-02T08:40:48.709771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:40:48.709771Z digest=sha256:0a799e606f6bd27df19e3673d248e8b4350afb0538eddc48bc8ea0abaecea6de