Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 50 inbound Pith citation observations for arXiv:2309.17179.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T11:47:31.242429Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation aea7076a-c96b-45b0-84c0-e62d2b2c97e4 · inbound
Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a9446a2e-c207-429b-a153-efdfed86db07 · inbound
Improve Mathematical Reasoning in Language Models by Automated Process Supervision Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cfbc782a-e2b0-4875-8f83-4673f0a6e7f6 · inbound
QLASS: Boosting Language Agent Inference via Q-Guided Stepwise Search Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdd185cc-189c-4585-a351-a8aeff4a8531 · inbound
On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ef57d7c-c201-481b-b7ab-80775d952152 · inbound
Policy Guided Tree Search for Enhanced LLM Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e92021b1-0c6e-4369-8774-ec43f7b0ccf9 · inbound
Bag of Tricks for Inference-time Computation of LLM Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb68f2ca-0621-4020-bd19-b463ba31ccd1 · inbound
Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fb995c4c-c156-4409-a4c3-748bfb0c287a · inbound
QM-ToT: A Medical Tree of Thoughts Reasoning Framework for Quantized Model Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation beb2dcd8-ce09-4a45-ac89-9df398fc6e18 · inbound
MMATH: A Multilingual Benchmark for Mathematical Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db427559-1d5c-4206-92ed-026d48eec8af · inbound
Can Past Experience Accelerate LLM Reasoning? Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8769437d-e1a6-40e5-a546-f22ebb8d114b · inbound
How Much Backtracking is Enough? Exploring the Interplay of SFT and RL in Enhancing LLM Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ad7cede-ca98-4c61-8f14-ecbbdd2bc5ab · inbound
Structured Pruning for Diverse Best-of-N Reasoning Optimization Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 650cc220-635d-40a7-87eb-790fc13501dd · inbound
Kinetics: Rethinking Test-Time Scaling Laws Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4234ff1d-a7b8-47fb-b9ec-4092ced0f56f · inbound
VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e67df85b-d344-4f51-8833-ae66fee42643 · inbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c54a79fa-7e5e-42a8-85a3-f7797b28d389 · inbound
TreeRL: LLM Reinforcement Learning with On-Policy Tree Search Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c47a8aa-1244-4b5e-bdc4-7b2deac56277 · inbound
KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6b7d3b68-72f8-48e2-94ec-d7f9494fb526 · inbound
Enhancing User Engagement in Socially-Driven Dialogue through Interactive LLM Alignments Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5eeb45b6-3c6d-4900-93a7-03e4147598f1 · inbound
Reasoning in machine vision by learning fast and slow thinking Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a1f5f4c-bd4b-4007-af55-873c9248df71 · inbound
Data Diversification Methods In Alignment Enhance Math Performance In LLMs Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1d90dd1-c705-42d0-baef-f00aca515744 · inbound
Enhancing Test-Time Scaling of Large Language Models with Hierarchical Retrieval-Augmented MCTS Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60d6e2b2-8357-4380-9b77-2c4b9d9fc6b8 · inbound
DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a39a040-6db1-4621-9b31-ed1e1b145f3a · inbound
It's Not That Simple. An Analysis of Simple Test-Time Scaling Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bacaa48-b2da-445a-9bf2-4aa70ba0a491 · inbound
LLM Economist: Large Population Models and Mechanism Design in Multi-Agent Generative Simulacra Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84845823-462d-4b71-adb7-056d11f47575 · inbound
Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14285457-87e1-4d05-a671-243df4b6e228 · inbound
Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 398b7170-af66-43e5-b8c6-f39c12d228aa · inbound
CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 1997
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f246477-bd08-4ba3-8922-04f6a7bc2519 · inbound
Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 028a9c3e-e00f-4664-bf40-d95bf6581261 · inbound
DiffCoT: Diffusion-styled Chain-of-Thought Reasoning in LLMs Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 788d1018-d4fc-4a17-8e50-9b4dfe58040a · inbound
Neural Chain-of-Thought Search: Searching the Optimal Reasoning Path to Enhance Large Language Models Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0754fcb6-bf24-483b-9a0b-0deb0bafddd8 · inbound
Agentic Reasoning for Large Language Models Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 120
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a5728366-d70a-4ffb-9660-d96556437e88 · inbound
Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a2df38d-1ab1-4573-b6f5-5d192859296b · inbound
Reinforce to Learn, Elect to Reason: A Dual Paradigm for Video Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 761dd4d6-29a3-4eb1-8972-b07c89f37a19 · inbound
Understanding Performance Gap Between Parallel and Sequential Sampling in Large Reasoning Models Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 66be9b4c-024e-473f-ae01-4a23529ff515 · inbound
Evaluation-driven Scaling for Scientific Discovery Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 93f3611e-068a-4ac9-b5a0-fb844a575fd3 · inbound
Distilling Long-CoT Reasoning through Collaborative Step-wise Multi-Teacher Decoding Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 64f389a3-ece9-4f54-8a47-654524236766 · inbound
SAT: Sequential Agent Tuning for Coordinator Free Plug and Play Multi-LLM Training with Monotonic Improvement Guarantees Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5210ee3d-d234-4f0a-a335-3a0f50e898ff · inbound
Maximizing Rollout Informativeness under a Fixed Budget: A Submodular View of Tree Search for Tool-Use Agentic Reinforcement Learning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ff54ab4a-a1f1-47c5-8e7f-8ebbb47b3abf · inbound
CA-SQL: Complexity-Aware Inference Time Reasoning for Text-to-SQL via Exploration and Compute Budget Allocation Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ada3a1b9-1026-41b1-9cad-17bf54fbcee6 · inbound
V-ABS: Action-Observer Driven Beam Search for Dynamic Visual Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cb9cf998-12a0-426d-826e-0ab8d015ecb7 · inbound
Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4f34326e-528b-4d96-9ad3-1fc96a76fa5d · inbound
Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4077066b-ce29-43c3-ad50-fc741a014a0d · inbound
Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9c889228-de92-44ce-852a-8375a79b6547 · inbound
Step-by-Step Optimization-like Reasoning in LLMs over Expanding Search Spaces Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1cc95fe5-48ec-40d7-879e-988120a46c37 · inbound
DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3c60c9d5-bffe-4fc4-937f-e71ab1969215 · inbound
TRACE: A Unified Rollout Budget Allocation Framework for Efficient Agentic Reinforcement Learning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e84f6b6a-14d9-492c-acf4-7991eabb59ea · inbound
APPO: Agentic Procedural Policy Optimization Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 956695dc-6775-41b4-930a-08ea6cbae14b · inbound
APPO: Agentic Procedural Policy Optimization Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26e4a92d-54d2-45ef-982a-7632f69be8e3 · inbound
Beyond Fixed Budgets: Characterizing the Inelasticity and Limitations of Tree-of-Thought Reasoning Strategies Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c83a8d58-2ff5-4cae-ac64-129a2da31a2c · inbound
DecompRL: Solving Harder Problems by Learning Modular Code Generation Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.