Pith. sign in

Paper Citation Record · LEDGER

Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2312.16682.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.16682 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T08:40:48.986418Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:29:16.737553Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bdd04f55-6a75-4367-b275-0d6bb227348b · inbound

Self-Rewarding Language Models cites this paper.

Self-Rewarding Language Models Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

Reference 121

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:01:42.518264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-13T12:01:42.290502Z digest=sha256:537fea74fb23a66852f31633b44956ce4b9fff4369645d463f1d23c17c1e76bc

Observation 3107b263-5fce-43a2-99a7-30228e15d883 · inbound

UNA: A Unified Supervised Framework for Efficient LLM Alignment Across Feedback Types cites this paper.

UNA: A Unified Supervised Framework for Efficient LLM Alignment Across Feedback Types Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-23T21:23:27.459223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T21:22:36.970101Z digest=sha256:acee1a2a47b93c878a382ab527f022e0bd985e9546c2f470ba56e20ece74d886

Observation 499568bc-e163-4943-a102-469d44a34e5d · inbound

Failure Modes of Maximum Entropy RLHF cites this paper.

Failure Modes of Maximum Entropy RLHF Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T14:02:39.774025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T14:02:11.084514Z digest=sha256:e5dea7616835a9cae7a3978274002baf1d0f8219dae36dc9e883e66193c0bdac

Observation 504c69b8-2128-4b3f-84ce-6f097940e55d · inbound

PoliLegalLM: A Technical Report on a Large Language Model for Political and Legal Affairs cites this paper.

PoliLegalLM: A Technical Report on a Large Language Model for Political and Legal Affairs Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:06:19.077085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T06:03:01.460352Z digest=sha256:0225fab021472366fae581e5b832c2dd9b00d95123122cac2a1be538c2e67b61

Observation 748042aa-09d4-4c81-afdf-ab8cd8118bff · inbound

Boosting Reinforcement Learning with Verifiable Rewards via Randomly Selected Few-Shot Guidance cites this paper.

Boosting Reinforcement Learning with Verifiable Rewards via Randomly Selected Few-Shot Guidance Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

Reference 89

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T03:19:42.952153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T03:18:26.590871Z digest=sha256:4e540872d007996b45c966eb34ca4102a73666430be737346f581d2163b76fe2

Observation b4ada445-124f-4b33-984a-81144f4444a0 · inbound

Self-Improvement Can Self-Regress: The Rise-and-Collapse Failure Mode of LLM Self-Training cites this paper.

Self-Improvement Can Self-Regress: The Rise-and-Collapse Failure Mode of LLM Self-Training Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T00:29:16.740668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T21:07:28.660119Z digest=sha256:739b3563f166fddf68e0f7380718baf3a74f3616f2397ad643275cd2a6b5a0c0

Observation c0cc8293-6a3a-4fef-8c49-44de339ed846 · inbound

World Feedback for Clinical Agents: Diagnosing RL in FHIR Environments cites this paper.

World Feedback for Clinical Agents: Diagnosing RL in FHIR Environments Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:18:55.876079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-07-03T20:14:46.147637Z digest=sha256:e2e8dca13ab8cacfff239ae94c30aa3e7e9e59a8ff3937a0f0386865b8065ad1

Observation 474341e2-d357-4988-9686-54c90fb4d7c6 · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

Reference 150

Resolution
unresolved
no resolver link, observed 2026-07-11T13:53:36.775836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T13:53:36.775836Z digest=sha256:9855d28babb56dc39cae2f562ec6c4ef8a230b3483da59047848e7d26d2b7da6

Observation 328acb1c-6e3a-4aee-b3d9-ce4f3de8e73f · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss

Reference 151

Resolution
unresolved
no resolver link, observed 2026-08-02T08:40:48.986418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:40:48.986418Z digest=sha256:3db051691f37c78fcf3b4a4dc24832b10cae0690944980f7fb06d151fcc000d1