Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:23:21.085081Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 2 inbound Pith citation observations for arXiv:2508.13070.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:23:21.085081Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T13:20:05.377960Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T13:20:05.555802Z
40 of 40 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7469ccae-3677-4bb7-87b5-911ffedba46c · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84f7b816-81c1-43d3-a1cd-0561871b180e · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86c64d48-c37f-4339-b76d-8c5cfd9b8312 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Gemini: A Family of Highly Capable Multimodal Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 298d0bbc-b325-420f-87da-ef90f1186c25 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Argmax flows and multinomial diffusion: Learning categorical distributions
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c905a289-a497-4ea0-80a7-5875453e0f27 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Structured denoising diffusion models in discrete state-spaces
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d13b7f15-35ac-4834-b368-1e01708ef52f · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Large Language Diffusion Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e47bdb06-d1c3-4c0e-a73e-51ef882f48e5 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Chain-of-thought prompting elicits reasoning in large language models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01175fc7-6e22-46e9-9b34-17927bc7b5c0 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning A survey on large language model based autonomous agents
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b09cb030-f1f0-4d84-a133-f51039183d36 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54bf7b2c-4eb5-43b7-8411-c679b471634c · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Causal language modeling can elicit search and reasoning capabilities on logic puzzles
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 965622e3-4e8f-4ce2-a5fc-254e3f6b7f85 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Beyond autoregression: Discrete diffusion for complex reasoning and planning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 05cf2003-d321-41ff-8fe1-09d15a573e3f · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Train for the Worst, Plan for the Best: Understanding Token Ordering in Masked Diffusions
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b497e3b-73ed-4ad7-a8e6-244403d68f4c · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Diffusion models: A comprehensive survey of methods and applications
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a211638-26ef-4dbb-8f77-6d104cb1f768 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning A survey on generative diffusion models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5ccadcb1-b9c3-4187-8069-e6163e26585e · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Autoregressive Models in Vision: A Survey
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abdb9721-fe84-40ae-b71d-0d37a25bb70e · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Language models are unsupervised multitask learners
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be4b17bb-5f69-4c0c-9cd8-52ac8caa2388 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Language models are few-shot learners
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0475881-861e-4b2e-9029-8b0282ac115a · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Denoising diffusion probabilistic models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f1f215f-056e-4198-b523-ae0bddb910a2 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Generative modeling by estimating gradients of the data distribution
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20286b7d-3b56-47e5-b2ea-dce477f2527b · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Score-based generative modeling through stochastic differential equations
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbb6b1e1-0dbd-40c8-be9d-d503b23bb464 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Xlnet: Generalized autoregressive pretraining for language understanding
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc348384-e46b-41f0-8dc1-baeae72b1fed · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Training and inference on any-order autoregres- sive models the right way
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8b4bec0d-bd87-41f6-a420-cd0c046bd30d · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Faith and fate: Limits of transformers on compositionality
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efa3cb39-95eb-419e-a77a-36cc8eb1636d · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning The pitfalls of next-token prediction
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 74239063-28d0-4d2d-a627-0f40d31a1492 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Autoregressive Modeling with Lookahead Attention
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation da8841ae-d26d-408e-bf5b-0a81161f3fce · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning On the planning abilities of large language models-a critical investigation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e6d8d945-1483-4d39-8439-bd2e859162b9 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Position: LLMs can’t plan, but can help planning in LLM-modulo frameworks
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3521d35e-cb3a-40ac-91ac-257051d4cad7 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Chain of thought empowers transformers to solve inherently serial problems
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8efeed51-41c7-433a-be23-f0476711b8ad · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Discrete diffusion modeling by estimating the ratios of the data distribution
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25fa6405-107a-44ac-a6a3-7aac8e209cd6 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning A theory of usable information under computational constraints
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 273a925d-a8fb-467e-8026-0f063e74dfd3 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning On the shortest arborescence of a directed graph.Scientia Sinica, 14:1396–1400, 1965
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation dcf4f9d5-7aa3-4bbe-8f44-ad6d3b4e0ab2 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Reinforcement learning: An introduction
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1083529e-3d01-4309-a996-3368cd70d481 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Reinforcement learning with deep energy-based policies
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af1b85f7-8cc2-437b-b162-a22ae9d4dfd7 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Proximal Policy Optimization Algorithms
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e29a70e9-5aa6-44de-869d-cde724d26a9b · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Lee, Kangwook Lee, and Dimitris Papailiopoulos
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 84f0489b-cafb-4187-8ced-3fe79644a78e · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Positional Description Matters for Transformers Arithmetic
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d965c38-77da-454e-8428-b4256bb5d65f · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Reverse That Number! Decoding Order Matters in Arithmetic Learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 2d4ad7ce-dee2-41de-add0-9c8c7bb0538a · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Transformers can do arithmetic with the right embeddings
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5f748538-feec-4615-9928-e9c37290ccb1 · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning A Closer Look at Invalid Action Masking in Policy Gradient Algorithms
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b5c6b90-a96b-41d2-91ab-c5a11eb91d7b · outbound
Reinforced Context Order Recovery for Adaptive Reasoning and Planning Focal loss for dense object detection
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 81703caa-0f45-45cb-aeae-62c064ed6958 · inbound
Any-Order Flexible Length Masked Diffusion Reinforced Context Order Recovery for Adaptive Reasoning and Planning
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 74d45714-c8de-4082-b695-fa2c9fc97465 · inbound
From Interface to Inference: Eliciting Any-Order Inference from Any-Order Models Reinforced Context Order Recovery for Adaptive Reasoning and Planning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.