Pith. sign in

Paper Citation Record · LEDGER

Clone-Robust AI Alignment

As of 19 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 4 inbound Pith citation observations for arXiv:2501.09254.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.09254 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:15:11.905544Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T22:50:33.207057Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:49:13.182463Z

Reference resolution

21 of 21 outbound references displayed

  • verified exact1
  • verified fuzzy2
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4a0ec520-0d0d-45c1-97ad-35ebf5e33ec5 · outbound

This paper cites Foundational Challenges in Assuring Alignment and Safety of Large Language Models.

Clone-Robust AI Alignment Foundational Challenges in Assuring Alignment and Safety of Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.764071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.764071Z digest=sha256:8bc5887985ad422314ec5c11fb93941332bc7b46e7c10e023ee848c8ac8b5ff7

Observation 8d4888a7-161e-4f2e-8048-a8ee1d2e01db · outbound

This paper cites Note that the function log er(x1) er(x1)+er(x2) is strictly convex in r(x1) and r(x2) as shown in Siththaranjan et al.

Clone-Robust AI Alignment Note that the function log er(x1) er(x1)+er(x2) is strictly convex in r(x1) and r(x2) as shown in Siththaranjan et al

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:12.269780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:15:11.905544Z digest=sha256:5326f6283b497eed70e967a293156683228f2ccf84c8d9ecf451b3c248e8c815

Observation 1579c556-f111-4291-9612-11d78047b2d0 · outbound

This paper cites MaxMin-RLHF: Alignment with Diverse Human Preferences.

Clone-Robust AI Alignment MaxMin-RLHF: Alignment with Diverse Human Preferences

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.788008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.788008Z digest=sha256:e0328dbc12d94d83be4d222c1a06cb4dda64483c6f8bb18441dcb4e2de2fe0ec

Observation 0e2e9106-73c6-480a-a06b-f762c21d181f · outbound

This paper cites Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback.

Clone-Robust AI Alignment Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.792641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.792641Z digest=sha256:c3ccf2b4e47131827a95048da46f7194bb498f635066034ed51a0c70e3f24adb

Observation 7be4ed5f-9d33-49e9-84a2-d4acb4b8c731 · outbound

This paper cites Mapping Social Choice Theory to RLHF.

Clone-Robust AI Alignment Mapping Social Choice Theory to RLHF

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.797365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.797365Z digest=sha256:a3a1e997ffc052ad08acf0a93fbabf7a75871e96bfcb7042af06b542f94dafbf

Observation a273bc8b-778e-4e97-b164-b19dbe6c9da2 · outbound

This paper cites an unresolved cited work.

Clone-Robust AI Alignment Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:12.347535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:15:11.814720Z digest=sha256:acfdf1807c058e9e1ed9080b42049fc3196df352087aa8031d27771e54f6665f

Observation 0b914654-272d-43a5-a17c-26fbff865f8f · outbound

This paper cites an unresolved cited work.

Clone-Robust AI Alignment Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-10T20:15:12.324347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:15:11.820274Z digest=sha256:eb174369aa8a4a8925f9f93259fdab0609b76c0b5310aaeb0aca4377d30b4aa4

Observation 1addc87a-f906-4f9c-b46f-643a8576225f · outbound

This paper cites Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning.

Clone-Robust AI Alignment Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.826805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.826805Z digest=sha256:f6bbdca73fe0e5cd7029f3e5e081726cb23881bb8d93d29d7bdd64cbdaaac75c

Observation 60d3818d-8a1a-445f-bb9a-ecb28acf6327 · outbound

This paper cites Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF.

Clone-Robust AI Alignment Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.833637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.833637Z digest=sha256:38d6c9f8ef75064b4f1effa54c126ebade01601bf18eafa28f16b3cfc952adcf

Observation 80d83fb8-77b6-408c-8023-f91cc88b43df · outbound

This paper cites A Roadmap to Pluralistic Alignment.

Clone-Robust AI Alignment A Roadmap to Pluralistic Alignment

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.839245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.839245Z digest=sha256:ced99d045db2754b063a22ca165095fb2756d3f3dfd4c7ebdf92ea244e8a7fac

Observation 1ce193b1-1f6d-411d-8c9e-e20b03b610f8 · outbound

This paper cites A Minimaximalist Approach to Reinforcement Learning from Human Feedback.

Clone-Robust AI Alignment A Minimaximalist Approach to Reinforcement Learning from Human Feedback

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.844164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.844164Z digest=sha256:924ebf63fef71c6da9dd80f932f056b9562a0b7d3d8a5368436347a4a3ee5243

Observation 5a84e697-4de5-473b-8299-cc0681639635 · outbound

This paper cites On the identifiability of mixtures of ranking models.

Clone-Robust AI Alignment On the identifiability of mixtures of ranking models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.854226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.854226Z digest=sha256:357833d5ce577ef1d8a1c291486a2adf9da1b5778a3b8391ea4244ac3147f045

Observation ead4f0c5-aa75-4294-afdc-61c37b25fd27 · outbound

This paper cites Provable Multi-Party Reinforcement Learning with Diverse Human Feedback.

Clone-Robust AI Alignment Provable Multi-Party Reinforcement Learning with Diverse Human Feedback

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.860439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.860439Z digest=sha256:8ffd4b217049d74cf393d1a63c6e539e9c3f17272c786f02d103391ff1523a5a

Observation 58808483-d97f-4e6c-9d95-f8da4d7ea6a8 · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

Clone-Robust AI Alignment Fine-Tuning Language Models from Human Preferences

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.866226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.866226Z digest=sha256:fd202747fbdb87ed8a9b19acec73e70691db6cd20bda17a0ddd9a82562c53569

Observation ad7ed927-ba3c-4554-b2b4-9bb1683bfeac · outbound

This paper cites Robust Reinforcement Learning from Corrupted Human Feedback.

Clone-Robust AI Alignment Robust Reinforcement Learning from Corrupted Human Feedback

Reference 1952

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.783150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.783150Z digest=sha256:a96470bd2ad5a7a8782c169e45e2d351aa1fa56029cdbcaa84e8a115386bd717

Observation c8821d90-dbfc-4d5e-bd97-b1fa53cf5ae3 · outbound

This paper cites RLHF and IIA: Perverse Incentives.

Clone-Robust AI Alignment RLHF and IIA: Perverse Incentives

Reference 1987

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.849157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.849157Z digest=sha256:22739065b7603e4212994f1a7f4a83f52aa6310d9c111595be13d44e52e98065

Observation 274973ab-cbae-4819-b5e4-62b3e0f81993 · outbound

This paper cites Axioms for AI Alignment from Human Feedback.

Clone-Robust AI Alignment Axioms for AI Alignment from Human Feedback

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.803558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.803558Z digest=sha256:71988ec658d035de5a5467102b43598480aa86a43987430eb051e1db0b7e34d6

Observation 15fee05c-a5c9-4b3a-ae2e-6c85836202ef · outbound

This paper cites Like us, Xu et al.

Clone-Robust AI Alignment Like us, Xu et al

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:15:12.294139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:15:11.874001Z digest=sha256:88af155047e315d9883cace86d4613e22c82514430b8d38643913393e38957c3

Observation 766b1459-b304-4282-8234-ec931e14f67e · outbound

This paper cites Obvious Independence of Clones.

Clone-Robust AI Alignment Obvious Independence of Clones

Reference 2022

Resolution
verified exact
local_arxiv, observed 2026-08-10T20:15:12.192329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T20:15:11.777165Z digest=sha256:35c832db3183f40c0caa8b339ad24de9c347d71335fa7ca2c5c7a1e9fd27ccfc

Observation 042204a1-9086-4a37-a54d-45d1a504af0c · outbound

This paper cites Corruption Robust Offline Reinforcement Learning with Human Feedback.

Clone-Robust AI Alignment Corruption Robust Offline Reinforcement Learning with Human Feedback

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.809668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.809668Z digest=sha256:d3945bae38bf94ccff3645a2d6bc0b7aaf8076f3c0ec4348fb062762573af5ae

Observation 02e70172-0199-43fc-b90b-06d4001b91e1 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Clone-Robust AI Alignment Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:11.770637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:15:11.770637Z digest=sha256:cc7132ea0a73cddb1895f5670af4cd61b991cf5040fbd3db32c3a087ab4285ae

Pith citing papers

Observation d30a5fa6-f7a2-4596-b4af-5a9f682e40df · inbound

Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers cites this paper.

Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers Clone-Robust AI Alignment

Reference 149

Resolution
unresolved
no resolver link, observed 2026-08-07T22:50:33.207057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T22:50:33.207057Z digest=sha256:c9fccfdaed2f4e6574267ff6e05dfb8df64ef466ff3af6004d652546b4e76a2f

Observation f9158f3a-d4a9-4615-8655-ff963d98d703 · inbound

Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? cites this paper.

Distortion of AI Alignment: Does Preference Optimization Optimize for Preferences? Clone-Robust AI Alignment

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:49:13.263112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T12:49:09.960375Z digest=sha256:a01758597ed32c6c9e7e53c74051ec0e2e3212e926551c1827ddf0e83466780f

Observation d89933d7-f66e-471e-b803-6fe0d74093e8 · inbound

Internal Pluralism and the Limits of Pairwise Comparisons cites this paper.

Internal Pluralism and the Limits of Pairwise Comparisons Clone-Robust AI Alignment

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-12T07:49:57.204875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T07:49:57.204875Z digest=sha256:fc0633189eb3ba153ed640de950d1f232440edd894924db85e40d6e4b856c5cc

Observation 716378cb-94ab-40ae-919d-bf78e599f517 · inbound

Internal Pluralism and the Limits of Pairwise Comparisons cites this paper.

Internal Pluralism and the Limits of Pairwise Comparisons Clone-Robust AI Alignment

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-02T09:03:14.618782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T09:03:14.618782Z digest=sha256:a3fafe7358afa3499f5b8973feeddb624febd98dbf7eb6d46bb9db536a0be8f9