Pith. sign in

Paper Citation Record · LEDGER

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback

As of 10 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2507.21131.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21131 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:14:16.727994Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 98fea892-9983-4f3c-bf58-55ae314d04c5 · outbound

This paper cites Concrete Problems in AI Safety.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Concrete Problems in AI Safety

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.642328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.642328Z digest=sha256:3f2b2ef00b58c57d87fc3ba8a49cffc147fc55b5048e44ca13b679242fa00fc0

Observation e5adc283-597b-4b00-82c4-e31932f82b47 · outbound

This paper cites AI safety via debate.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback AI safety via debate

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.663769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.663769Z digest=sha256:f2ac2e1c082ceb16ca8e6c57336e20ea7d56705b8a3a47d65877ab8d118aa51e

Observation 211cf569-56ac-4ca5-b6fb-5609d1d9d85e · outbound

This paper cites Language Models (Mostly) Know What They Know.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Language Models (Mostly) Know What They Know

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.669363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.669363Z digest=sha256:81cac341713bc5a13002e2f714d5fc99fa88fce6215914a32fa604420e1452fe

Observation 19245a6b-857e-4a22-b0fa-cae434bd318c · outbound

This paper cites Algorithms for inverse reinforcement learn- ing.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Algorithms for inverse reinforcement learn- ing

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:17.071285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:14:16.684460Z digest=sha256:56c9fab96c47bb21a1141b267ecb7aefadc396ccadd20659d48476d6cf700256

Observation 436d7b01-70a3-4706-abef-90a5af555369 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.693748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.693748Z digest=sha256:06c327e12596f4606319230c54ec5ea446cd578f3b678ef9c634a587686b3d32

Observation 06b5a3b5-ecfd-448d-91dd-f902350515fc · outbound

This paper cites A new system-wide diversity measure for recommendations with efficient algorithms.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback A new system-wide diversity measure for recommendations with efficient algorithms

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T15:14:16.828662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:14:16.703719Z digest=sha256:32c88225e088da48e67523986f584795c0d30d82827ee20ca07dbc865e746515

Observation 0fe0c28a-d246-4698-861e-37cba15abb66 · outbound

This paper cites The Fates of Merging Supermassive Black Holes and a Proposal for a New Class of X-Ray Sources.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback The Fates of Merging Supermassive Black Holes and a Proposal for a New Class of X-Ray Sources

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T15:14:16.803484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:14:16.708646Z digest=sha256:d7c864ae89f3b04c74ffd8c4a5027776fdf5d6a3e1b66894f163e444095482d1

Observation 02197dd9-d428-456a-b78c-87ad2378f729 · outbound

This paper cites Red Button.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Red Button

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:17.034115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:14:16.727994Z digest=sha256:db026d7c67acd3cb6d391cee435269ff628a2d86505d2af52bf0559bf86178a0

Observation 8a1a899e-d56f-4157-9daa-0fef7608adf1 · outbound

This paper cites Discovering Latent Knowledge in Language Models Without Supervision.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Discovering Latent Knowledge in Language Models Without Supervision

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.689102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.689102Z digest=sha256:67ee70f801a1de30301267022e04eeca10fc71966b51149985fab4bc3b97aa27

Observation dd39c946-5aac-4e79-aa74-8da336e6e031 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.648067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.648067Z digest=sha256:2a1a15413a3473f2e25744f8bb062e780fdec9fb99efc2d3fd8cf166a3f719df

Observation ad369e9c-2031-4c53-8021-f4544014d216 · outbound

This paper cites Supervising strong learners by amplifying weak experts.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Supervising strong learners by amplifying weak experts

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.652974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.652974Z digest=sha256:6c295c3ae548b7681864199193ff8188470323df383bff3ac45c77068ac18d58

Observation a1494bd8-4baf-4426-8e33-f0d8c6450a55 · outbound

This paper cites Improving alignment of dialogue agents via targeted human judgements.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Improving alignment of dialogue agents via targeted human judgements

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.657949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.657949Z digest=sha256:4ada9b4d9a1814aeb5bc5469b581547eca4987e70cd7d41da8ad69a965294683

Observation 4cbba20d-fc0e-45f2-833b-801d7b7fbe1b · outbound

This paper cites Self-critiquing models for assisting human evaluators.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Self-critiquing models for assisting human evaluators

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.698792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.698792Z digest=sha256:77c5a5ad70b532686f9df1d27bd212695bd419b34ad355e8628a5643f15b9aa7

Observation 858244ef-3017-4256-819a-15683edcdaf7 · outbound

This paper cites Spin-Spin Coupling at Small $x$: Worm-Gear and Pretzelosity TMDs.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Spin-Spin Coupling at Small $x$: Worm-Gear and Pretzelosity TMDs

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.713331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.713331Z digest=sha256:3bf92467f3ee29a7035210aafd96e97e13268c80be53a5db1370823751ab3455

Observation d9c4c7f2-95ab-46a5-925c-e1e001df020f · outbound

This paper cites Teaching language models to support answers with verified quotes.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Teaching language models to support answers with verified quotes

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:16.679333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:16.679333Z digest=sha256:15304e237d428629175e20db986a424280ddae1cedc560d10f118d0545618e5b

Observation 0719c205-ff88-44cc-a248-71620dc3f487 · outbound

This paper cites Pebble: Feedback- efficient interactive reinforcement learning via relabeling experience and un- supervised pre-training.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Pebble: Feedback- efficient interactive reinforcement learning via relabeling experience and un- supervised pre-training

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:17.087493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:14:16.674730Z digest=sha256:ef7627ffa5b35a0d75b20b0cb0782733ba6d3172d6cd0edb127b16413ac1f7d5

Observation 8029b33b-822d-4a61-b8c2-f69458cf70f9 · outbound

This paper cites Red button,.

NPO: Learning Alignment and Meta-Alignment through Structured Human Feedback Red button,

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:17.052660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T15:14:16.720775Z digest=sha256:8928be6ba6b5a738df338130718a3190ffef0ea0f692f8449380e026543b95f2

Pith citing papers

No inbound Pith citation observations are available.