Pith. sign in

Paper Citation Record · LEDGER

A Critical Evaluation of AI Feedback for Aligning Large Language Models

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2402.12366.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.12366 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:13:39.934107Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T14:16:23.883268Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 77c7ca05-d268-4b24-b2ae-d5b0ad946c8c · inbound

Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters cites this paper.

Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters A Critical Evaluation of AI Feedback for Aligning Large Language Models

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:16:23.884831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T14:16:23.842092Z digest=sha256:8bb9de954b8d76edfbd64c046fc140792a15808a800ca90c676237800dbc1abf

Observation 4264866e-441e-4e9a-9651-e1a6ca68c512 · inbound

Refusal Tokens: A Simple Way to Calibrate Refusals in Large Language Models cites this paper.

Refusal Tokens: A Simple Way to Calibrate Refusals in Large Language Models A Critical Evaluation of AI Feedback for Aligning Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T19:25:04.860028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T19:25:04.860028Z digest=sha256:3454b940c30824e21a024e4ce503f2cea2756a5e772a7e7444245ccadc2241b6

Observation 90462ffc-cea6-4706-9af7-c9bdb6ec8da8 · inbound

Data-Centric Improvements for Enhancing Multi-Modal Understanding in Spoken Conversation Modeling cites this paper.

Data-Centric Improvements for Enhancing Multi-Modal Understanding in Spoken Conversation Modeling A Critical Evaluation of AI Feedback for Aligning Large Language Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T10:57:01.361372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T10:57:01.361372Z digest=sha256:1527d11211c541d33df2d7d2e2711e4247307db748a6504cd90cc2d781391573

Observation 0229ae0e-ac7c-46c8-b5f9-6683334a4ce5 · inbound

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges cites this paper.

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges A Critical Evaluation of AI Feedback for Aligning Large Language Models

Reference 181

Resolution
unresolved
no resolver link, observed 2026-08-06T14:13:06.611360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:13:06.611360Z digest=sha256:b24129fccd4f2c4c2c0f80b18b6bb5d40897d8cb31eac50909f78ce4074ec69c

Observation da3c88fb-2c0e-4778-b16a-bec6e054be78 · inbound

Graph-Enhanced Policy Optimization in LLM Agent Training cites this paper.

Graph-Enhanced Policy Optimization in LLM Agent Training A Critical Evaluation of AI Feedback for Aligning Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T07:17:44.647263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:17:44.647263Z digest=sha256:466b1e609566815b19b526622e6a3ad080367f2688253bc5881b6a0c7a0a28ba

Observation 0bd1725e-c5cb-4fc4-8e4c-984832f445b6 · inbound

Toward a Theory of Value in AI Alignment cites this paper.

Toward a Theory of Value in AI Alignment A Critical Evaluation of AI Feedback for Aligning Large Language Models

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-14T04:13:39.934107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:13:39.934107Z digest=sha256:3552bd8944b9468ce4f2a817b96105ea5b5be087b440183fba54dd320025e0d2