Pith. sign in

Paper Citation Record · LEDGER

Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2402.10958.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.10958 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T16:18:40.567531Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4ad7b4d5-f69b-4d43-a61b-cdfbbbb65288 · inbound

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment cites this paper.

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-24T00:38:39.836064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-24T00:38:36.992597Z digest=sha256:a0c2818794e20f394d1f29e63d1379aa280b188cc03237745613c0f391b26774

Observation 649dace3-01ac-4ac4-8cca-b3a0e6853cb3 · inbound

On Almost Surely Safe Alignment of Large Language Models at Inference-Time cites this paper.

On Almost Surely Safe Alignment of Large Language Models at Inference-Time Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T16:18:40.567531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:18:40.567531Z digest=sha256:63329aa9eb28a6c5e31a6a333d249dad85b4369615607bba0ddcc47b6f5c3f16

Observation b6ce621d-1846-4d13-b133-bd353dcd43ae · inbound

Reviving The Classics: Active Reward Modeling in Large Language Model Alignment cites this paper.

Reviving The Classics: Active Reward Modeling in Large Language Model Alignment Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T11:47:17.652900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:47:17.652900Z digest=sha256:bb6681639c45e23bf9319ad51fd692d86086cdb72a11bda9cb6ce09614b9b5e9

Observation 2a9e79ab-9ede-473a-910b-1ae2daf2d4d9 · inbound

ChartSketcher: Reasoning with Multimodal Feedback and Reflection for Chart Understanding cites this paper.

ChartSketcher: Reasoning with Multimodal Feedback and Reflection for Chart Understanding Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:56.873323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:56.873323Z digest=sha256:507bfc1a07d43ab09d9306c7415c45d5c386fa88446e918739df8b4b006f28de

Observation ac69e7f0-10ed-4d9b-8ad0-9b06eba55910 · inbound

Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities cites this paper.

Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-06T16:34:25.272362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:34:25.272362Z digest=sha256:8148d27c964a5864bae0d4c942e9bdf24942c65c34684aa35db1c744a8512a62

Observation 3202da34-0f25-437e-b9e8-d4835183269e · inbound

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects cites this paper.

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:12:30.106598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-18T08:11:31.181704Z digest=sha256:c6f6402ddaea5cbe2bc27fe3567e8d2f8759123e576bb1c5ababd32ccd73ae7a

Observation 77500ca3-58fd-49d1-8ecc-9497e8e1e968 · inbound

RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection cites this paper.

RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:36:35.294034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-15T20:33:09.627731Z digest=sha256:d9402044dc18b35a43c2fdb048a883752b1bc0d2210123b988888b2f1535ff5c

Observation 89595ea6-5a5b-4048-83d3-84593ecc7695 · inbound

Generating Place-Based Compromises Between Two Points of View cites this paper.

Generating Place-Based Compromises Between Two Points of View Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:01:12.278619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-08T03:36:31.695964Z digest=sha256:95071a2c8e4f9a828fa237855c4646a6e0f21de72c827009e61dfc22152985f5

Observation 73bcca5e-ca59-4b78-a8ec-22f726ff617d · inbound

Beyond Pairs: Your Language Model is Secretly Optimizing a Preference Graph cites this paper.

Beyond Pairs: Your Language Model is Secretly Optimizing a Preference Graph Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:30:58.410493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-11T02:26:33.933264Z digest=sha256:d237e97e9e02ad6697758bdce2d82637989a2ed28ab2c47782244adad9de742a

Observation 4d9973fc-054e-46c5-b440-63a5e531ee52 · inbound

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation cites this paper.

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-06-30T08:34:26.735555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-30T08:30:37.016856Z digest=sha256:655e79b6c9b449691a00cf321ba1ef051511a20aa1e72ad30aabd9c2621cc3ee