Pith. sign in

Paper Citation Record · LEDGER

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO

As of 8 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 3 inbound Pith citation observations for arXiv:2505.15694.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15694 v1

Coverage vector

measured 27 of 27 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:19:09.839089Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:45:05.223275Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T20:13:13.616657Z

Reference resolution

27 of 27 outbound references displayed

  • verified exact3
  • verified fuzzy9
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cf07d1dc-5adf-4e5b-92bb-781c88a6afce · outbound

This paper cites Claim E.4.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Claim E.4

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:19:11.083198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:09.754235Z digest=sha256:801c49c3d8251c4dce058d06670ddde3b29ae2c165e536a45bf75410eb7c264c

Observation 4f432d9e-58b8-45dc-9942-58deb2a88a92 · outbound

This paper cites Thus, by Lemma H.2, we have with probability at least 1 − δ, 1 n nX i=1 ηixi 2 ≤ C · σ · r 1 + ln(1/δ) n , for some universal constant C >0.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Thus, by Lemma H.2, we have with probability at least 1 − δ, 1 n nX i=1 ηixi 2 ≤ C · σ · r 1 + ln(1/δ) n , for some universal constant C >0

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:19:10.976972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:09.788906Z digest=sha256:7b0147499ef390e1323d89ebeec04d8abe238e0f02aab6b15498537c5618aa20

Observation 8ac5c9d9-abc4-42c5-9969-35103f8b9218 · outbound

This paper cites Finally, we use the win rate from these comparisons as our primary performance metric, following the methodology outlined in the DPO paper (Rafailov et al., 2023).

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Finally, we use the win rate from these comparisons as our primary performance metric, following the methodology outlined in the DPO paper (Rafailov et al., 2023)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:19:11.194972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:09.705227Z digest=sha256:79de510ce4160b1856b7fee203cdbed36ae37b59bd9a90d32d0e7316756a833c

Observation 8aa47b56-c5b2-4a2c-ab47-3ca7a5fa914b · outbound

This paper cites Manipulation attacks in local differential privacy.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Manipulation attacks in local differential privacy

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:19:11.868113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:08.590930Z digest=sha256:d591e874c8229b0ec047edb626472e4d59900f3f029b7a54f5a203efc207f15c

Observation eb00ceec-6840-4f45-91ae-7ba1ebbef8c7 · outbound

This paper cites Differentially Private Reward Estimation with Preference Feedback.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Differentially Private Reward Estimation with Preference Feedback

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T15:19:10.315913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:08.710420Z digest=sha256:4e3154542f83d57a233d663248e74cad97795112ec84bf3e762444946ea5bf98

Observation 6f6217b9-2b21-4366-b350-b0d52c2f1c40 · outbound

This paper cites Provably Robust DPO: Aligning Language Models with Noisy Feedback.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Provably Robust DPO: Aligning Language Models with Noisy Feedback

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:08.819245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:08.819245Z digest=sha256:178babe69fbda73f4b65cc1574ef3ee43741ef59d4bca96deb9b125d416431fb

Observation 03059725-2090-47c3-94f8-d5d389713b53 · outbound

This paper cites C., Jordan, M.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO C., Jordan, M

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:19:11.734979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:08.905079Z digest=sha256:b93b8bdb21cb44558308be301e1f195f0db9ec13c3bca417646891004ce9011f

Observation 4e27a73e-4b82-4766-831b-99727d532408 · outbound

This paper cites A tail inequality for quadratic forms of subgaussian random vectors.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO A tail inequality for quadratic forms of subgaussian random vectors

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:09.094011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:09.094011Z digest=sha256:3a0d305a8ae78210c059a704bc501a915cc7bf27718bb3c4a67e2c3ce469953a

Observation f79755dd-1d76-499c-877f-e45cedcb5d0d · outbound

This paper cites Corruption Robust Offline Reinforcement Learning with Human Feedback.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Corruption Robust Offline Reinforcement Learning with Human Feedback

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:09.300383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:09.300383Z digest=sha256:e3e1c30fa56ecfb36347df33d807705d21204809be20ecd505a7e5317b3b0784

Observation d9904036-a23e-4ef2-b25c-ac1113470fd4 · outbound

This paper cites Dueling RL: Reinforcement Learning with Trajectory Preferences.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Dueling RL: Reinforcement Learning with Trajectory Preferences

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:09.343297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:09.343297Z digest=sha256:ab712402790349b51a338c9e19a95b22838fe732244494c1fbe4cf330473cc30

Observation 7187fcf9-6b79-4ced-b4d4-3c4c4aa06903 · outbound

This paper cites The importance of online data: Understanding prefer- ence fine-tuning via coverage.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO The importance of online data: Understanding prefer- ence fine-tuning via coverage

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:19:11.617733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:09.372738Z digest=sha256:8ab4d95d953781fa42c3e89437bed0e771ec8cb63aaa90eeb6c2f914a7d386a8

Observation 4c1e7082-13de-40a7-a7a9-ffe05b95f5db · outbound

This paper cites Provable Offline Preference-Based Reinforcement Learning.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Provable Offline Preference-Based Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:09.500623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:09.500623Z digest=sha256:3a6efb6ce0cc08f6d9b774074578a341a8a97a4927b03a8e038feb5bab9a7f2e

Observation e6bcd209-91de-4231-b8dd-d47083cb7456 · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Fine-Tuning Language Models from Human Preferences

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:09.536408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:09.536408Z digest=sha256:093b7a72c1c1ff2c46692966ce11f4e65bdb4a31694c7d8554888636c7101017

Observation 8342ec0c-c248-43a5-81ee-c13432b97416 · outbound

This paper cites Additional Related Work We discuss here more relevant work that do not fit in the main text.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Additional Related Work We discuss here more relevant work that do not fit in the main text

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:19:11.507737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:09.579591Z digest=sha256:4c5eb24cc7f7302ac4b092e8193211581e23446f2cb6898f58ac215a6bbf12d1

Observation 7cbe31bb-7cdd-491b-b353-54ac960f42ad · outbound

This paper cites an unresolved cited work.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:19:11.408244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:09.625588Z digest=sha256:af65f42cf6093d1b53c5e8bf592e58a68c6b8f17dd93b73550f71fdca836972e

Observation c7fd88bc-2068-45e5-8c2b-5d0552f1d9b5 · outbound

This paper cites rejected.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO rejected

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:19:11.299828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:09.660226Z digest=sha256:fe62c584275c723082fca88c8049bcb4d396ad7f20369922710b91b6e217a6bb

Observation a8e615f1-ea4c-4250-8402-d0df3edebdc4 · outbound

This paper cites Chosen” and “Rejected.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Chosen” and “Rejected

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:19:10.866476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:09.839089Z digest=sha256:0f6a71ea7dea36f1c8702790e6d6f7a8821e59e8acb6ae202dbd7d0734f293af

Observation 791c3347-9199-4ea2-90c8-b9eb91a052ad · outbound

This paper cites Robust Reinforcement Learning from Corrupted Human Feedback.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Robust Reinforcement Learning from Corrupted Human Feedback

Reference 1952

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:19:10.582859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:08.152208Z digest=sha256:d75ef223819b26309ed6c99c5008e7bd18526b2a48ca993428a7c4a044d633f5

Observation 7b7c3549-e50f-4f13-b26a-6ea1eb994b06 · outbound

This paper cites Exploratory Preference Optimization: Harnessing Implicit Q*-Approximation for Sample-Efficient RLHF.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Exploratory Preference Optimization: Harnessing Implicit Q*-Approximation for Sample-Efficient RLHF

Reference 1965

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:09.445181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:09.445181Z digest=sha256:f284c6c1ee01fb9a1a7d7ed201907fa2c270c1c3e8267af4976f9d92c7f83678

Observation 3e81486b-232c-4b7b-b93b-ba130a1d6d1d · outbound

This paper cites Correcting the Mythos of KL-Regularization: Direct Alignment without Overoptimization via Chi-Squared Preference Optimization.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Correcting the Mythos of KL-Regularization: Direct Alignment without Overoptimization via Chi-Squared Preference Optimization

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:09.183433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:09.183433Z digest=sha256:c7583bde6535e482976f4a453e9ed5248580102902ccdf47796603868b88375b

Observation e8c9eb8f-a6e9-4b2c-8181-a1e88602ff0c · outbound

This paper cites Exposing Privacy Gaps: Membership Inference Attack on Preference Data for LLM Alignment.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Exposing Privacy Gaps: Membership Inference Attack on Preference Data for LLM Alignment

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:09.010064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:09.010064Z digest=sha256:50351579c755ed952be8905a83a85251453091e820f6bdde2d6895c20bc4f7d5

Observation f5aa4cf2-f486-4ab7-a49d-81279ea92799 · outbound

This paper cites Classification Under Misspecification: Halfspaces, Generalized Linear Models, and Connections to Evolvability.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Classification Under Misspecification: Halfspaces, Generalized Linear Models, and Connections to Evolvability

Reference 2019

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:19:10.432203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:08.491110Z digest=sha256:7a2ab3ba197365801d6f1069d31defd93eefeb28635faa74daaa401637f82651

Observation b539ee6f-cebb-4b93-93e9-627b6d143fdb · outbound

This paper cites Reinforcement Learning for LLM Post-Training: A Survey.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Reinforcement Learning for LLM Post-Training: A Survey

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:09.406012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:09.406012Z digest=sha256:d88f865a58845dcf7c3bd92459dbcf3e4700d9f31869c0d02da6a289efc91097

Observation 40ae400d-0164-464d-9486-e1460f501434 · outbound

This paper cites Trimmed Maximum Likelihood Estimation for Robust Learning in Generalized Linear Models.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Trimmed Maximum Likelihood Estimation for Robust Learning in Generalized Linear Models

Reference 2021

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:19:10.757471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:19:07.927499Z digest=sha256:b9a8b3eafd0a4ca73cd50195945093e49d9ea1b81314ddc925bc8536ed6ef7f5

Observation 63cf1714-d86f-4816-a113-9bd108e73088 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:08.018639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:08.018639Z digest=sha256:24467b1a9b9bb01ede4929faa3e1c3aae4b39fce593a5929f41d22119609cb66

Observation 9c8f5672-e659-495c-8784-2e003c483aad · outbound

This paper cites Value-Incentivized Preference Optimization: A Unified Approach to Online and Offline RLHF.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Value-Incentivized Preference Optimization: A Unified Approach to Online and Offline RLHF

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:08.360025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:08.360025Z digest=sha256:e8ca3eba9de1e1b6896835b0cc421dd9e7ebbfb978dd58bbcd2de9460ac450ee

Observation 9bf3d758-8c50-40ea-88e2-3edf4d488e02 · outbound

This paper cites Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback.

A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T15:19:08.255598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:19:08.255598Z digest=sha256:9a8d3b3930ab72c583412b5de75ea94007738e89e98df1641b55b8be45c16ad3

Pith citing papers

Observation f1837fc3-a321-403a-b96b-5aad2be5b011 · inbound

Square$\chi$PO: Differentially Private and Robust $\chi^2$-Preference Optimization in Offline Direct Alignment cites this paper.

Square$\chi$PO: Differentially Private and Robust $\chi^2$-Preference Optimization in Offline Direct Alignment A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:05.223275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:45:05.223275Z digest=sha256:a8ba5efb274c8d6a45f0ccf06df3ddcfb0f7cfa89a27c606d536dda87cacd874

Observation cd1778d0-8b58-41b0-928e-ab3f08ae8001 · inbound

Reinforcement Learning from Human Feedback: A Statistical Perspective cites this paper.

Reinforcement Learning from Human Feedback: A Statistical Perspective A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:13:13.620120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T20:10:43.578904Z digest=sha256:5a6d0601b60b7ce08109c403618f4c9feb8eda399cc9c7b0b95a42de3fcc102f

Observation 0157b7c1-6d1c-441f-87e2-2170d262c79c · inbound

Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training cites this paper.

Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T06:32:24.252477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T06:30:51.812541Z digest=sha256:7df23f477917bdfb2a47ba93f42a84929fd8ebfdc16a56c7b1110e7d7d2bd9a1