Pith. sign in

Paper Citation Record · LEDGER

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment

As of 13 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2411.10534.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.10534 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T19:39:52.676052Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 69793b3b-3e62-4f68-a6ac-631df9c59531 · outbound

This paper cites Beyond Preferences in AI Alignment.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Beyond Preferences in AI Alignment

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.595975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.595975Z digest=sha256:d6aeee757e9a55119744d19236a4dddf25bb4289a86d6ed943eaf81dc456db2b

Observation e683f59b-4f15-4a41-b53b-5d14e70e1558 · outbound

This paper cites Deliberative Technology for Alignment.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Deliberative Technology for Alignment

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.600574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.600574Z digest=sha256:cf900a041b8a92a6838189642a20e7dd0bf8d48eecb40b2fdf33154376c54b26

Observation b2c0cebe-2234-4b17-9b8f-eb0c71e25f36 · outbound

This paper cites Manning, and Chelsea Finn.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Manning, and Chelsea Finn

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.604691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.604691Z digest=sha256:a3cc6aaab77e0726214ac6c5828b4830003d97edab93a5074abade44f1f9f596

Observation dbca8042-4ca2-457e-8409-dbe1128d72f1 · outbound

This paper cites Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Is DPO Superior to PPO for LLM Alignment? A Comprehensive Study

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.612999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.612999Z digest=sha256:dd27516590fb9fb56ed833c1533c6ea1430d67f3226aaaf7dd3e7d3d74a24f1f

Observation e5c260d2-ff12-4dda-896f-2cba2f51396e · outbound

This paper cites Direct preference-based policy optimization without reward modeling.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Direct preference-based policy optimization without reward modeling

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:39:52.966308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T19:39:52.617004Z digest=sha256:4f6a32f42dfb7a1cddc060e7120b9fe9fd5b8e9ba38ef77ef9121c8e490a6048

Observation 4f4fdb30-2753-4bae-b614-7210da193576 · outbound

This paper cites Beyond Reverse KL: Generalizing Direct Preference Optimization with Diverse Divergence Constraints.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Beyond Reverse KL: Generalizing Direct Preference Optimization with Diverse Divergence Constraints

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.620622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.620622Z digest=sha256:d81276e3e7cb6c2e81de45cc68e92772abb6612bb15414397aa36f6b4cd83961

Observation b507da81-d45b-4eaa-b2c8-9f6a244236d1 · outbound

This paper cites an unresolved cited work.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-12T19:39:52.955176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T19:39:52.624164Z digest=sha256:a2eb526909ef167532a6461e99100396b047a91809708b3b9ad2f8b8495e4bf3

Observation 03eb4060-2812-4518-9ed5-1bf46321e9ea · outbound

This paper cites Learning to summarize with human feedback.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Learning to summarize with human feedback

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.627591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.627591Z digest=sha256:2756a4094ebb9a431b104e140561274d616ab760293fbedfc369da4a9f53dcff

Observation 27a118fc-0d2c-40a1-a6d2-45b105be6539 · outbound

This paper cites Deep reinforcement learning from human preferences.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Deep reinforcement learning from human preferences

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.630421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.630421Z digest=sha256:b528dbe69dd780d451abcb103095d76e1508d7853fe5254cd6c7c8d62885e29e

Observation 109095c4-f7d5-435a-a3e9-b3c8894cb8db · outbound

This paper cites Improving alignment of dialogue agents via targeted human judgements.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Improving alignment of dialogue agents via targeted human judgements

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.633285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.633285Z digest=sha256:f2107272989bad336c0c24dfa25f12775a9d59443df4885710748c722c889bd1

Observation 41a79c77-0549-4c67-b638-67f13c1c4dd3 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.636422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.636422Z digest=sha256:4ac96e04547d3484e1a1ca5075f8e56ef79b3dd1c819c6afc06e82c20fb843ca

Observation 58982c2f-1cd4-4ff3-867a-2eee19425b61 · outbound

This paper cites Jury learning: Integrating dissenting voices into machine learning models.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Jury learning: Integrating dissenting voices into machine learning models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:39:52.930031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T19:39:52.640437Z digest=sha256:d5f068fbcec8c614de79dc70c1b872bff9042be912afe383f0ad0d85adee777b

Observation 6931099f-7b9a-4544-acea-90ba83baaa35 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Constitutional AI: Harmlessness from AI Feedback

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.643877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.643877Z digest=sha256:36f4f1afba6a9ce89f914b581e9c8a75ec4302bb7c0a743c4177a7d9380a01cd

Observation 9196ab73-006f-42b9-a3b6-0872b8104ab6 · outbound

This paper cites Liao, Esin Durmus, Alex Tamkin, and Deep Ganguli.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Liao, Esin Durmus, Alex Tamkin, and Deep Ganguli

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.647474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.647474Z digest=sha256:18a3590c81cefde49463a1b8f0b3d7db0027f90ad4cc19a4086c9aea2daa0c8d

Observation 590043db-3736-4e0c-a7c8-ac12182dae2d · outbound

This paper cites Rule- based rewards for language model safety, 2024.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Rule- based rewards for language model safety, 2024

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:39:52.918756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T19:39:52.650573Z digest=sha256:692a1243a009fcef67784eaede34a1c65302a2374c34100f7b9d470ae482adee

Observation f4469d5c-a084-4ae8-a1cc-d61c59f13394 · outbound

This paper cites Specific versus General Principles for Constitutional AI.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Specific versus General Principles for Constitutional AI

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.653860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.653860Z digest=sha256:3d5459b21979c2cf81c57bf648a228495afc723bfd1d96c376b913db3c572535

Observation 8f3849b7-91c4-4410-88bb-3776415f1b7e · outbound

This paper cites Rule-based reinforcement learning for efficient robot navigation with space reduction.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Rule-based reinforcement learning for efficient robot navigation with space reduction

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:39:52.907710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T19:39:52.657688Z digest=sha256:aeece7dea7336f027b29186dd99a8d84cb2931bdfd44674c5efeb0657f5e0b6e

Observation a1956bef-af8c-4402-b8f9-560006877d8f · outbound

This paper cites Democratic Policy Development using Collective Dialogues and AI.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Democratic Policy Development using Collective Dialogues and AI

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.661399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.661399Z digest=sha256:5da6f01730dfcdcf0b94054d3b7b8c6b22d57dae1ea49db7ef99345768b38a75

Observation ca1a902e-34d3-46ed-9c2e-90d8e603b608 · outbound

This paper cites Inverse Constitutional AI: Compressing Preferences into Principles.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Inverse Constitutional AI: Compressing Preferences into Principles

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.665058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.665058Z digest=sha256:db8b8a317b7226bd7b3a9536b821ce473d9c84c199d0961185183ec661f800f1

Observation e859dcb8-59b1-4005-bc0c-84ab48944b0b · outbound

This paper cites Qiu, Michael Varga, and Aviv Ovadya.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Qiu, Michael Varga, and Aviv Ovadya

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:39:52.895654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T19:39:52.668925Z digest=sha256:4ee74c225309a4d2f82f3d72b7cca10cbfdff0068e6569eac4fff0dfc3d05742

Observation d49ab429-8548-4a11-b8f1-c3c01680bdba · outbound

This paper cites Mémoire sur les élections au scrutin.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Mémoire sur les élections au scrutin

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T19:39:52.884725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T19:39:52.672404Z digest=sha256:bb00851dac3c38cc4bef71bb20840432eaff4daeafe8697b1b6d8415b5f59301

Observation 68ead250-29e0-41b8-b584-055f85b030b8 · outbound

This paper cites super-majority.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment super-majority

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.676052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.676052Z digest=sha256:314808308f8d595d43d0300b30f3f3a014f8b469916caef967a62192338a6d42

Observation 748ec504-d1ab-41de-90ee-2339fd5f3b81 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T19:39:52.608618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:39:52.608618Z digest=sha256:d8be65fd263a3e34fb3d0931f4854d03f9882dcde81299f3ca14f9850e202e9a

Pith citing papers

No inbound Pith citation observations are available.