Pith. sign in

Paper Citation Record · LEDGER

Generating Sequences by Learning to Self-Correct

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 25 inbound Pith citation observations for arXiv:2211.00053.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2211.00053 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 25 of 25 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:33:37.690667Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

30
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a4cfd5eb-0d23-4567-aa9a-5e63c8736292 · inbound

Language Models can Solve Computer Tasks cites this paper.

Language Models can Solve Computer Tasks Generating Sequences by Learning to Self-Correct

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-17T12:17:26.672384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T12:17:26.602361Z digest=sha256:c9d93b8426407c23b863b8010f5dd8397ac8be3915fb6f1ca91be44ff0915f68

Observation 1af583fc-6f28-4d99-8a6e-4474319cb138 · inbound

Self-Refine: Iterative Refinement with Self-Feedback cites this paper.

Self-Refine: Iterative Refinement with Self-Feedback Generating Sequences by Learning to Self-Correct

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-10T20:47:39.644362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T20:47:39.476572Z digest=sha256:70266b00dcaeda75af0157d00d8d65fda7d6003565aeccf5d784efe4260e8507

Observation 99dd202d-51ed-48c4-b3e8-60caf322d5e7 · inbound

Reasoning with Language Model is Planning with World Model cites this paper.

Reasoning with Language Model is Planning with World Model Generating Sequences by Learning to Self-Correct

Reference 96

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T01:49:29.053361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T01:49:28.796581Z digest=sha256:9c29b3595594369e5d7260d8c694ed6ef49fd28938bf9360b8593401e967eb18

Observation 7017aa68-a473-489d-aab9-151c251fe455 · inbound

LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code cites this paper.

LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code Generating Sequences by Learning to Self-Correct

Reference 141

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T17:34:42.714345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T17:34:42.565806Z digest=sha256:5e046e5c674048f4a974f73d3194745ff62c3ae72d535b8dd5b6b9b392f556d2

Observation c1ccffb0-b099-420a-8576-85c3fde1634f · inbound

S$^2$-MAD: Breaking the Token Barrier to Enhance Multi-Agent Debate Efficiency cites this paper.

S$^2$-MAD: Breaking the Token Barrier to Enhance Multi-Agent Debate Efficiency Generating Sequences by Learning to Self-Correct

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T21:33:37.690667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T21:33:37.690667Z digest=sha256:721fcd3ed171d9ee47133bae05e488fb29b7b63e358d7654f86a3d394188ea6a

Observation f2538324-2c16-4852-9bbb-ca936096e965 · inbound

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection cites this paper.

ReflectEvo: Improving Meta Introspection of Small LLMs by Learning Self-Reflection Generating Sequences by Learning to Self-Correct

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:40.628932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:40.628932Z digest=sha256:20109ff83a649db32688613397080b2a847a924397eca86f43a184cbfcc0521b

Observation 2d8a13ba-1438-4563-baba-31d72530b596 · inbound

Boosting LLM Reasoning via Spontaneous Self-Correction cites this paper.

Boosting LLM Reasoning via Spontaneous Self-Correction Generating Sequences by Learning to Self-Correct

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:51:30.726142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:51:30.726142Z digest=sha256:91718e096da9125d4dd182d081fc9eb840df22bf4c3ad35e98dad16377a61146

Observation 9440cc8e-c0e8-44ac-9c32-05538e2e1dc5 · inbound

Hallucination Detection with Small Language Models cites this paper.

Hallucination Detection with Small Language Models Generating Sequences by Learning to Self-Correct

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T23:11:43.449396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:11:43.449396Z digest=sha256:31415e9b1e57f5b9d3cd491c8a3ced3deea568e412b53966512f68e2abbec04c

Observation 9bf4c6f0-5e06-425b-a958-c3657d0419fd · inbound

R4ec: A Reasoning, Reflection, and Refinement Framework for Recommendation Systems cites this paper.

R4ec: A Reasoning, Reflection, and Refinement Framework for Recommendation Systems Generating Sequences by Learning to Self-Correct

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T14:58:27.811196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:58:27.811196Z digest=sha256:7f88e89cd5ebb758050ae999cad9e4509ef506356d1c57466aa44cc5a7da34fc

Observation b4ff55bf-e67c-40b1-aaba-fca55260873d · inbound

I2CR: Intra- and Inter-modal Collaborative Reflections for Multimodal Entity Linking cites this paper.

I2CR: Intra- and Inter-modal Collaborative Reflections for Multimodal Entity Linking Generating Sequences by Learning to Self-Correct

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T05:11:13.226381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:11:13.226381Z digest=sha256:bd6c885ef19e9813aac3acf2594a3e7ca74a4006de41ea6a29ebf35a04b1001e

Observation a0a86057-9ea6-4940-a317-98ab80e2f66f · inbound

SC-Captioner: Improving Image Captioning with Self-Correction by Reinforcement Learning cites this paper.

SC-Captioner: Improving Image Captioning with Self-Correction by Reinforcement Learning Generating Sequences by Learning to Self-Correct

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T22:58:39.580049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:58:39.580049Z digest=sha256:403cdfbf3c83706bf6ca35118e74a2ef4a1f1378e1b5f6c0b1cbc56c789fbb1a

Observation dff38631-8e78-4ff6-9b04-94c90dd22b76 · inbound

User-Assistant Bias in LLMs cites this paper.

User-Assistant Bias in LLMs Generating Sequences by Learning to Self-Correct

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:51:53.168403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T22:51:02.400926Z digest=sha256:951963ff26d674b4ff19292e3ff3585eae1e7b9b8ec37b1ef067c82cf050e601

Observation bf968da1-1b58-4f92-ae0f-23c5652b6afd · inbound

Don't Act Blindly: Robust GUI Automation via Action-Effect Verification and Self-Correction cites this paper.

Don't Act Blindly: Robust GUI Automation via Action-Effect Verification and Self-Correction Generating Sequences by Learning to Self-Correct

Reference 5

Resolution
malformed identifier
arxiv_id, observed 2026-05-10T23:15:48.775192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:15:27.813769Z digest=sha256:b31dbb0b2738ee5ea7ca1255410fe7c8475ac42559f2b2635e5cdcd3ac21d6b9

Observation 4ec7c14c-25f2-4f9b-898f-be71da5dc299 · inbound

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding cites this paper.

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding Generating Sequences by Learning to Self-Correct

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:16:06.221861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T23:05:05.251150Z digest=sha256:6fb8ea9de7b318492459c23ed405e238e80247d66fcfef8dec7c26b9be561e62

Observation 3144e4ad-54fa-4146-a574-e58fd832317e · inbound

Perturbation Dose Responses in Recursive LLM Loops: Raw Switching, Stochastic Floors, and Persistent Escape under Append, Replace, and Dialog Updates cites this paper.

Perturbation Dose Responses in Recursive LLM Loops: Raw Switching, Stochastic Floors, and Persistent Escape under Append, Replace, and Dialog Updates Generating Sequences by Learning to Self-Correct

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:55:31.788886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T19:17:06.375875Z digest=sha256:fe9e3d544fa3484f18a9d4818ed6b0f7c76538576808ae0ced54c562089b5873

Observation a4072c17-5f75-4c47-aa50-65a93279cf73 · inbound

STRIDE: Learnable Stepwise Language Feedback for LLM Reasoning cites this paper.

STRIDE: Learnable Stepwise Language Feedback for LLM Reasoning Generating Sequences by Learning to Self-Correct

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:49:00.664961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T20:47:16.236629Z digest=sha256:f6a7204fc4d365a2e02c5abd76f1980cd494f488557a5b7e0df2bcec605caab0

Observation 3d350414-44ee-4180-90d8-5023b5a2ee6a · inbound

EPPC-OASIS: Ontology-Aware Adaptation and Structured Inference Refinement for Electronic Patient-Provider Communication Mining in Secure Messages cites this paper.

EPPC-OASIS: Ontology-Aware Adaptation and Structured Inference Refinement for Electronic Patient-Provider Communication Mining in Secure Messages Generating Sequences by Learning to Self-Correct

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:04:52.533247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T15:54:58.741719Z digest=sha256:d9bd02b817345b2e20ad671684edb44b4d36a249365664d474889bfcd944199c

Observation 7d992c4d-6284-4b5b-b08d-3b4105e8a474 · inbound

DenoiseRL: Bootstrapping Reasoning Models to Recover from Noisy Prefixes cites this paper.

DenoiseRL: Bootstrapping Reasoning Models to Recover from Noisy Prefixes Generating Sequences by Learning to Self-Correct

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-29T11:53:23.676256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T11:51:01.769882Z digest=sha256:4bc6bb75f38311afd6480da90d201fb07d9cc160952875c58e05c10a0d3a53ac

Observation 2403ae88-3b05-4e42-b867-308952f29aa0 · inbound

DenoiseRL: Bootstrapping Reasoning Models to Recover from Noisy Prefixes cites this paper.

DenoiseRL: Bootstrapping Reasoning Models to Recover from Noisy Prefixes Generating Sequences by Learning to Self-Correct

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T12:58:10.959787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:58:10.959787Z digest=sha256:a5f505809d0311c9b7c7f6d430e128dfda1920621d89d0c850d97109eaa24aab

Observation 79402cc3-55ce-4fe6-aa86-cea42fe63330 · inbound

DRIFT: Decoupled Rollouts and Importance-Weighted Fine-Tuning for Efficient Multi-Turn Optimization cites this paper.

DRIFT: Decoupled Rollouts and Importance-Weighted Fine-Tuning for Efficient Multi-Turn Optimization Generating Sequences by Learning to Self-Correct

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:02:50.291825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T23:16:49.358792Z digest=sha256:f5d2587c8f12d9e0107bc1e343223796556882edb7201e8f624b07a42cbfb2c1

Observation 2bf46bd7-37f4-4d3b-8eec-5218b6e1f930 · inbound

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination cites this paper.

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination Generating Sequences by Learning to Self-Correct

Reference 96

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:23.947729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T19:58:32.016341Z digest=sha256:4613b7d6384e5225f2164ac1c3985f8879cbd78466d4c3e02f9e450054750723

Observation e688db68-11f6-4d51-91df-5653e8e69e1c · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Generating Sequences by Learning to Self-Correct

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-11T13:53:36.775836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T13:53:36.775836Z digest=sha256:1717a2b9b2027e1eb0530295596dc1d5870e4ef467c9207a946bd53298f47574

Observation 24b67dc4-17e7-4270-8a97-7b57822cb507 · inbound

Multi-Turn On-Policy Distillation with Prefix Replay cites this paper.

Multi-Turn On-Policy Distillation with Prefix Replay Generating Sequences by Learning to Self-Correct

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T08:40:34.391896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:40:34.391896Z digest=sha256:3450380043b8332914f3a1f021f69d6d31874dcab545e7368596fc2facca13d8

Observation 76c6f394-d6e1-4557-8a99-8089f7dcf0cb · inbound

ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning cites this paper.

ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning Generating Sequences by Learning to Self-Correct

Reference 50

Resolution
malformed identifier
no resolver link, observed 2026-08-05T04:53:13.165423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:53:13.165423Z digest=sha256:327a557bf60dc6124d670c75a0d0886c687c56e2b6ab9f23f56db77513f868fa

Observation 146908c8-e1e7-4a67-bb45-aa5792d2b4ae · inbound

Refining Over Resampling: Test-Time Self-Correction for LLM Reasoning cites this paper.

Refining Over Resampling: Test-Time Self-Correction for LLM Reasoning Generating Sequences by Learning to Self-Correct

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T05:05:24.740924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T05:05:24.740924Z digest=sha256:e67eabed16d1dc0489bbb841a0647ba712b0f1929f2b7a4121dc8ca68b6bb909