Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2607.07492.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 04efd97e-4c87-4693-b3da-09fefec4eec1 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning When is tree search useful for LLM planning? it depends on the discriminator
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a16fe72f-3fb7-4006-8bab-092a40233808 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 528fae0e-a982-4944-9f53-70f1a8d68583 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Stream of Search (SoS): Learning to Search in Language
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8c2bb982-28d4-484d-a387-3f1bb133d009 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 712ec4db-f69d-44e6-b7b8-23d371193dda · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Reasoning with language model is planning with world model
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 39ec3e35-f6c6-4910-8042-1819539f62d3 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Large Language Models Cannot Self-Correct Reasoning Yet
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f633b95b-6d56-452b-bbda-536be2d83a10 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning When can LLMs actually correct their own mistakes? A critical survey of self-correction of LLMs
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 28e80778-acf0-44fc-8fb5-1a0dab638af4 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5e4dd129-b271-4f5a-907d-f1278642043e · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Training Language Models to Self-Correct via Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e6f72bb4-d634-4582-b432-7cf52f4f9d92 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Beyond A*: Better Planning with Transformers via Search Dynamics Bootstrapping
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f3afc76f-deaf-4123-898a-688eec3811dd · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Self-Refine: Iterative Refinement with Self-Feedback
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b1eb2a18-cfed-4797-8be5-27b3bd8b044f · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Recursive Introspection: Teaching Language Model Agents How to Self-Improve
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation febfa471-d05b-48ab-abe9-dd8d53fcd2fa · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Qin, T., Alvarez-Melis, D., Jelassi, S., and Malach, E
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 07f73784-95e0-4cbd-8408-e75d5948032c · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Spurious Rewards: Rethinking Training Signals in RLVR
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 98b87332-a28c-4333-8ba7-ddb08456fe8f · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning From Reasoning to Super-Intelligence: A Search-Theoretic Perspective
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f3adb9db-3608-400c-8937-ba8c2c1783d1 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Reflexion: Language Agents with Verbal Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f8a8195d-db19-4b5c-977a-b3aa8e89fe20 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Dualformer: Controllable Fast and Slow Thinking by Learning with Randomized Reasoning Traces
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8d8d45b-d57b-4936-ad1d-e84a32ad4dce · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 99e4dd5d-e16b-42c9-b22a-a741252bfbe9 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8ec762a6-6cd1-4d8f-869c-aa544ee39612 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Step Back to Leap Forward: Self-Backtracking for Boosting Reasoning of Language Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4f060ae9-68da-414f-b53b-e3e3177154f7 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Tree of Thoughts: Deliberate Problem Solving with Large Language Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ffb1fdfd-b29b-429a-a97f-2cfa26290d4c · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3260f571-67dd-4554-845e-87cabc304fec · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning How Much Backtracking is Enough? Exploring the Interplay of SFT and RL in Enhancing LLM Reasoning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1df115b3-2ae8-47a6-b2a7-1af672e9dc49 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning ASTRO: Teaching Language Models to Reason by Reflecting and Backtracking In-Context
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 29b89142-652b-45e5-aa93-9bb432f2a561 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Beyond markovian: Reflective exploration via bayes-adaptive rl for llm reasoning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c9fc3acb-9ac7-484e-abf1-17309c514171 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e020ec10-504d-4913-a94c-091ce6bde7df · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d6a6729-a5ca-414c-aa49-15f52a989fe1 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning PlanBench: An Extensible Benchmark for Evaluating Large Language Models on Planning and Reasoning about Change
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5f69c0ca-ffff-49a8-bb4b-8aa467a6eecb · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning LLMs Can't Plan, But Can Help Planning in LLM-Modulo Frameworks
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9f1d35d0-ab26-4ac4-8d09-224773124057 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e9d07a35-be41-449c-85cd-8224346feb51 · outbound
Search, Fail, Recover: A Training Framework for Correction-Aware Reasoning Proceedings of the 16th
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.