Pith. sign in

Paper Citation Record · LEDGER

Chain of Hindsight Aligns Language Models with Feedback

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2302.02676.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2302.02676 v8

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:50:24.592291Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

27
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b45d4549-9528-4c2c-affd-c1858f77b22e · inbound

Aligning Text-to-Image Models using Human Feedback cites this paper.

Aligning Text-to-Image Models using Human Feedback Chain of Hindsight Aligns Language Models with Feedback

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:39:15.486004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-14T20:39:15.388881Z digest=sha256:b16f354337b92626b33757bcebe9b035fd00132623cc66ee8d60982e0554d606

Observation df7e02a7-8c44-4344-a897-c31194eb80f5 · inbound

Teaching Large Language Models to Self-Debug cites this paper.

Teaching Large Language Models to Self-Debug Chain of Hindsight Aligns Language Models with Feedback

Reference 108

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:24:24.960444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-12T06:24:24.607354Z digest=sha256:c846b20e956ee5594f86c740e29c71da188cd19fa7d13922e2f6de15f3fef8b3

Observation 21d65289-9823-4462-98ee-d142f83cceea · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models Chain of Hindsight Aligns Language Models with Feedback

Reference 172

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:39.552064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:f75bd1fb3feaec71ecd46ef289ca6907555a02d3b1c9483b43510626e952dd58

Observation 888cf924-ce4e-4e37-a18d-8028d9859137 · inbound

The Rise and Potential of Large Language Model Based Agents: A Survey cites this paper.

The Rise and Potential of Large Language Model Based Agents: A Survey Chain of Hindsight Aligns Language Models with Feedback

Reference 127

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:47:47.954838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-11T10:47:44.152066Z digest=sha256:e7c3b82e4f497e81556128605473fc3e869be560ad570ef810418e68ffef1ec9

Observation 78fe38be-df0c-4001-aa54-18a1e7f11068 · inbound

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation cites this paper.

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation Chain of Hindsight Aligns Language Models with Feedback

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:00:51.564458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T22:00:51.487120Z digest=sha256:49497d4b3d0a2f3af7e2aafb9fe01425e28dc8bbc722f3b55831a004ae834a6a

Observation 51f31bb9-f395-427e-8558-ebef725cee31 · inbound

TrustLLM: Trustworthiness in Large Language Models cites this paper.

TrustLLM: Trustworthiness in Large Language Models Chain of Hindsight Aligns Language Models with Feedback

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.494101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:029be2a89a04c4af121a064d0efd495bc24d1e98aa82885c3fde3b146aa4ea0a

Observation f059d631-c20d-490d-bdab-2df8213cf7b3 · inbound

LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code cites this paper.

LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code Chain of Hindsight Aligns Language Models with Feedback

Reference 142

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T17:34:42.717565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-10T17:34:42.565806Z digest=sha256:3d30c2ac36d65b75f4843b81b8036bbca346d9c670b83f443f217751e64d68da

Observation b82aee17-ec1b-44d3-96ce-91fa6d62a8f4 · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Chain of Hindsight Aligns Language Models with Feedback

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:26.442731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:48692e5ded42c495f7811eca60217a18c7ea4e8ff323f6da8c717245c634b4ff

Observation b34c4e40-7eed-4521-8663-53c83a10bd00 · inbound

A Survey on LLM-as-a-Judge cites this paper.

A Survey on LLM-as-a-Judge Chain of Hindsight Aligns Language Models with Feedback

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:35:44.001994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-23T17:33:13.394338Z digest=sha256:39e881bbb30b8f44269699803eebdab8db0506ca2d03768ef16b9a2f6b9962ac

Observation 0ff99883-5672-4257-aa20-22714700c191 · inbound

Compressed Chain of Thought: Efficient Reasoning Through Dense Representations cites this paper.

Compressed Chain of Thought: Efficient Reasoning Through Dense Representations Chain of Hindsight Aligns Language Models with Feedback

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T04:47:40.422808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-17T04:47:40.335477Z digest=sha256:0c2a023669842934bd69a2ab48c7b4ee816aefdafa81b55b9cb52d0cb9e38179

Observation fb952474-2ed1-46da-a955-4ebbabf38e69 · inbound

Creating an LLM-based AI-agent: A high-level methodology towards enhancing LLMs with APIs cites this paper.

Creating an LLM-based AI-agent: A high-level methodology towards enhancing LLMs with APIs Chain of Hindsight Aligns Language Models with Feedback

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T13:38:04.398514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:38:04.398514Z digest=sha256:54d95045056e3a7801db7f31cf5212ae235f88e3203a104dbf877aed0fe52715

Observation 58de8c53-5f0d-41a4-8fe0-f01c3065712c · inbound

RapGuard: Safeguarding Multimodal Large Language Models via Rationale-aware Defensive Prompting cites this paper.

RapGuard: Safeguarding Multimodal Large Language Models via Rationale-aware Defensive Prompting Chain of Hindsight Aligns Language Models with Feedback

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T04:29:04.272864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:29:04.272864Z digest=sha256:505c314583cc75c20366fe93889e006a93abb998992197fb7238de8f00f5952d

Observation 3b56215b-5b92-4e90-8d5c-c405bcaa78fe · inbound

LLM-Powered Multi-Agent System for Automated Crypto Portfolio Management cites this paper.

LLM-Powered Multi-Agent System for Automated Crypto Portfolio Management Chain of Hindsight Aligns Language Models with Feedback

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T22:48:26.304765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:48:26.304765Z digest=sha256:6ce58f6732513c948c91b74f11b40a29bdeaa21cd16b3c08e106c8a48339cb31

Observation baa0cbb8-ab8a-4928-9b07-4f8f16043630 · inbound

A Comprehensive Survey of Agents for Computer Use: Foundations, Challenges, and Future Directions cites this paper.

A Comprehensive Survey of Agents for Computer Use: Foundations, Challenges, and Future Directions Chain of Hindsight Aligns Language Models with Feedback

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-23T05:02:35.944176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-23T04:59:36.994758Z digest=sha256:bae76b437e9cc49c0e25cbf6f49d7d6c0145248cc6da8f4b0b509ac0fd86b732

Observation ee38643e-f9ed-4dd6-8425-bf363f386b1e · inbound

Automatically Generating Rules of Malicious Software Packages via Large Language Model cites this paper.

Automatically Generating Rules of Malicious Software Packages via Large Language Model Chain of Hindsight Aligns Language Models with Feedback

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T10:50:24.592291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:50:24.592291Z digest=sha256:25d6e5f7610448f02ab55b75dd3da17f4464b2f3a5d1b1f64b779a407588f42c

Observation 2289331b-09d8-49c6-9224-51191d8bc78d · inbound

Generative to Agentic AI: Survey, Conceptualization, and Challenges cites this paper.

Generative to Agentic AI: Survey, Conceptualization, and Challenges Chain of Hindsight Aligns Language Models with Feedback

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-16T10:10:25.014226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:10:25.014226Z digest=sha256:ef595570bb359160836953c305082171ee7e138d80b8d88e8500f1e79ca90dbd

Observation e7c2b313-9bb8-48d7-872c-9e44e93999a4 · inbound

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time cites this paper.

Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time Chain of Hindsight Aligns Language Models with Feedback

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:20.272639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:20.272639Z digest=sha256:b4e6045973db1b2b6463ad7b2807522f73e38545059b1ad365511589450c8974

Observation 1fc1ce79-7efe-4436-8321-0ee83f34fb0f · inbound

Large Language Model-Based Agents for Automated Research Reproducibility: An Exploratory Study in Alzheimer's Disease cites this paper.

Large Language Model-Based Agents for Automated Research Reproducibility: An Exploratory Study in Alzheimer's Disease Chain of Hindsight Aligns Language Models with Feedback

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:02:49.235474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:02:49.235474Z digest=sha256:dcf271eed1718431c13983a9d88df8fdd1275de49d619c2302133cca62e9efe2

Observation 3b9a7e87-1844-403c-9371-9a48e56b292e · inbound

Multi-level Value Alignment in Agentic AI Systems: Survey and Perspectives cites this paper.

Multi-level Value Alignment in Agentic AI Systems: Survey and Perspectives Chain of Hindsight Aligns Language Models with Feedback

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:04.446080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:04.446080Z digest=sha256:0bcd405284f38934f5ddacaa60d581a9eaa0484f9416d870ecd193bb5573cbac

Observation 63beea3b-8356-4f0c-96e8-868abdd8ef81 · inbound

From Outcomes to Processes: Guiding PRM Learning from ORM for Inference-Time Alignment cites this paper.

From Outcomes to Processes: Guiding PRM Learning from ORM for Inference-Time Alignment Chain of Hindsight Aligns Language Models with Feedback

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:57:56.730028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:57:56.730028Z digest=sha256:e9c252c4767e0c50b08e72ec7cff46a23e068de6aad62b968cc096f9e69df21b

Observation c9492d50-6ce5-4b04-bdcf-c9d856e340f1 · inbound

Invariant-based Robust Weights Watermark for Large Language Models cites this paper.

Invariant-based Robust Weights Watermark for Large Language Models Chain of Hindsight Aligns Language Models with Feedback

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T18:33:07.896741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:33:07.896741Z digest=sha256:8abba42bf075f679fc42d0dd30457056bed72677d72ea0e0f35017f05f57931c

Observation 84a467ab-def9-44bc-904a-b457e82c56f9 · inbound

Feedback-Driven Execution for LLM-Based Binary Analysis cites this paper.

Feedback-Driven Execution for LLM-Based Binary Analysis Chain of Hindsight Aligns Language Models with Feedback

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:44:37.914449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T10:40:32.133423Z digest=sha256:4c82067cf8e7ef9b79b97005bd023bce8e76b2ac5afb9a03eca5c2318f4f474c

Observation 9dd5bbb0-c187-4884-a236-952f529115b9 · inbound

Learning from Language Feedback via Variational Policy Distillation cites this paper.

Learning from Language Feedback via Variational Policy Distillation Chain of Hindsight Aligns Language Models with Feedback

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:39:00.296891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T20:34:36.764090Z digest=sha256:c2f0cce1a0f83b9eb32046a8de0f15b408fba6e3f4a4f6e2ed05dbca1a64ea25

Observation 4c312351-8c69-4ad3-a390-d8dfa220f5ac · inbound

Reinforcing Human Behavior Simulation via Verbal Feedback cites this paper.

Reinforcing Human Behavior Simulation via Verbal Feedback Chain of Hindsight Aligns Language Models with Feedback

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:24:02.456916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-21T07:21:48.649289Z digest=sha256:aea654078fa8ea7694ab6e3abfc8e672c3e4eb421708647e71e56d6293ff06f4

Observation 007cf121-c003-41fe-861a-e03429beac57 · inbound

AI as a Tool for Simulation-Based Experiments in Literary Studies cites this paper.

AI as a Tool for Simulation-Based Experiments in Literary Studies Chain of Hindsight Aligns Language Models with Feedback

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-28T15:02:19.091390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-28T14:54:14.909995Z digest=sha256:ee4325e866a14ee4ec461d527864b7648b5a6bfdcd683808e6e19defcf8b7922

Observation 6c2bb689-008d-4bad-8790-6691437e2c51 · inbound

LeAct: Learning to Reason from Expert Actions cites this paper.

LeAct: Learning to Reason from Expert Actions Chain of Hindsight Aligns Language Models with Feedback

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T06:32:20.842111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:32:20.842111Z digest=sha256:52115ec369ebe728039c72a57fa90613519540ddcb30237e1f147de53bd61e94

Observation 85eac940-cd36-4287-b81c-573dad37b8fa · inbound

Training with (Swap) Regret Loss in a Single-Layer Self-Attention Model: A Case Study on the Probability Simplex cites this paper.

Training with (Swap) Regret Loss in a Single-Layer Self-Attention Model: A Case Study on the Probability Simplex Chain of Hindsight Aligns Language Models with Feedback

Reference 102

Resolution
unresolved
no resolver link, observed 2026-07-31T23:51:58.137701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T23:51:58.137701Z digest=sha256:8d5f3ec4d95131b06d30e2dbdca68f317a5a4f98d11962cee5af07f9a80d7226

Observation 20999843-4ea5-49a2-a9bb-a9e225dff3ff · inbound

Different Feedback, Different Updates: Selective Self-Learning from User Interactions for Large Language Models cites this paper.

Different Feedback, Different Updates: Selective Self-Learning from User Interactions for Large Language Models Chain of Hindsight Aligns Language Models with Feedback

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T23:27:32.658523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:27:32.658523Z digest=sha256:f1cb87309dff62bc9be23908fb4eb8c75ffddc67bf76352f87569079cf5981d6

Observation 57c82f25-9014-4a54-83da-572a149951a1 · inbound

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling cites this paper.

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling Chain of Hindsight Aligns Language Models with Feedback

Reference 202

Resolution
unresolved
no resolver link, observed 2026-08-12T14:10:45.897030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:10:45.897030Z digest=sha256:d3908ddc60988875fb7a0dc1214382148c7c35464012787a69323fb6d6749c27