Pith. sign in

Paper Citation Record · LEDGER

Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2402.10958.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.10958 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:25:56.873323Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4ad7b4d5-f69b-4d43-a61b-cdfbbbb65288 · inbound

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment cites this paper.

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-24T00:38:39.836064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-24T00:38:36.992597Z digest=sha256:2df33c7983ed761221b111c145ab4746501770da9132c5119797db84d6c4bd17

Observation 2a9e79ab-9ede-473a-910b-1ae2daf2d4d9 · inbound

ChartSketcher: Reasoning with Multimodal Feedback and Reflection for Chart Understanding cites this paper.

ChartSketcher: Reasoning with Multimodal Feedback and Reflection for Chart Understanding Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:56.873323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:56.873323Z digest=sha256:a3cbfddcb66a4596b63f73c747ec5fe9d97f0579e5fe4350ab4d8ffc3d297548

Observation ac69e7f0-10ed-4d9b-8ad0-9b06eba55910 · inbound

Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities cites this paper.

Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-06T16:34:25.272362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:34:25.272362Z digest=sha256:cca658977d9ab765a5a52f45c0e9487cdbf8a325094b5509faf0adecccddf7e9

Observation 3202da34-0f25-437e-b9e8-d4835183269e · inbound

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects cites this paper.

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:12:30.106598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T08:11:31.181704Z digest=sha256:b179b442d94bd10ee4e34ac093adefe3ca23184da17c7cc75a0998984f55911e

Observation 77500ca3-58fd-49d1-8ecc-9497e8e1e968 · inbound

RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection cites this paper.

RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:36:35.294034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T20:33:09.627731Z digest=sha256:04d7db5d278d211946b56b8c6798b1a56b91d6a7998fccf7e46d48bc3d0dd769

Observation 89595ea6-5a5b-4048-83d3-84593ecc7695 · inbound

Generating Place-Based Compromises Between Two Points of View cites this paper.

Generating Place-Based Compromises Between Two Points of View Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:01:12.278619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T03:36:31.695964Z digest=sha256:6ea09a451cd60e002cb4ec1473148b86d9d05034c333ec4265bd80377efaf99e

Observation 73bcca5e-ca59-4b78-a8ec-22f726ff617d · inbound

Beyond Pairs: Your Language Model is Secretly Optimizing a Preference Graph cites this paper.

Beyond Pairs: Your Language Model is Secretly Optimizing a Preference Graph Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:30:58.410493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T02:26:33.933264Z digest=sha256:aaa07a140aac245262fade871b807cfcd2cf35b7f3b28957403e77e1d6c93525

Observation 4d9973fc-054e-46c5-b440-63a5e531ee52 · inbound

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation cites this paper.

Reward-Free Code Alignment from Pretrained or Fine-Tuned LLM: Unpacking the Trade-offs for Code Generation Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-06-30T08:34:26.735555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T08:30:37.016856Z digest=sha256:35a19335541bf51967ee075d0ab5c2eb5ffe4cdf13fabba65d2915da3173e8c4