Pith. sign in

Paper Citation Record · LEDGER

Enhancing LLM Reasoning with Reward-guided Tree Search

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2411.11694.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.11694 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:06:00.892822Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b747b821-6732-41e9-9d67-8e74890b5ad4 · inbound

Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems cites this paper.

Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:35:31.441347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T00:35:31.375020Z digest=sha256:98d49150b3f1269404e335cd456f0fb11b89aab45b483b1336c0b84c04eaaeea

Observation 6aa1724b-52d3-41c5-8e25-14eb70d34815 · inbound

Search-o1: Agentic Search-Enhanced Large Reasoning Models cites this paper.

Search-o1: Agentic Search-Enhanced Large Reasoning Models Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:36:27.714815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T17:36:27.515468Z digest=sha256:117eec7ed94215356d5f0742452a5ccc7b87d1e25b6fa4b3c1f9a69dba5bc593

Observation 0826ecf3-81b6-4e94-976d-9fe0cd154d7f · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 228

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.527788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:9d75bedd93360728d7ef3ff4bb6d9cb911f049e09abef8e4013583bc2c8f0c76

Observation 4731a328-5df5-481d-91b0-31ab3a55478f · inbound

Large Language Model Agent: A Survey on Methodology, Applications and Challenges cites this paper.

Large Language Model Agent: A Survey on Methodology, Applications and Challenges Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 108

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:52:10.570222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:51:34.309870Z digest=sha256:2ef3bbfca3299f1d73aaec75c3ef4327520a1019c3193dbbbf09edb6485a49db

Observation 34a2bdca-22ac-4241-bad8-498eb7451526 · inbound

ZeroSearch: Incentivize the Search Capability of LLMs without Searching cites this paper.

ZeroSearch: Incentivize the Search Capability of LLMs without Searching Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-17T17:44:13.470477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T17:44:13.310155Z digest=sha256:d615abb60110afd9ffeecbe0e1fde16b1603688e5f8490d93a44263326ed12ae

Observation 56fe818b-eff3-4366-8fa8-a5ab4045d2de · inbound

ZeroSearch: Incentivize the Search Capability of LLMs without Searching cites this paper.

ZeroSearch: Incentivize the Search Capability of LLMs without Searching Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-22T16:06:45.978606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T16:05:04.715678Z digest=sha256:5833c2d038c3953503e6bb25fc7a5267e6b6065f3647878f704fa09556f8dec4

Observation a80ca779-edbd-4c3a-80ff-78034a3c8734 · inbound

EquivPruner: Boosting Efficiency and Quality in LLM-Based Search via Action Pruning cites this paper.

EquivPruner: Boosting Efficiency and Quality in LLM-Based Search via Action Pruning Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:06:00.892822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:06:00.892822Z digest=sha256:5728db112ea0e7d7cc899083840bb29ec7bc94197f7c77e8f98571c393268f7f

Observation 9f9ffe20-ea0b-4f37-9e95-257e32a45eae · inbound

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning cites this paper.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.406362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.406362Z digest=sha256:c31864e897fed90460de5818e5ceb2c390585205a3a171e352c3b8ce44ad1652

Observation f9b003d7-32fe-4b32-ab05-8497d578fab5 · inbound

Reward Model Generalization for Compute-Aware Test-Time Reasoning cites this paper.

Reward Model Generalization for Compute-Aware Test-Time Reasoning Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:27.309575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:42:27.309575Z digest=sha256:32dfdcb42b59c3ac15caee0eb97c903cb6790e023aff0f45e119346d0d594ca3

Observation c4faad56-4c78-4937-9050-740d21861537 · inbound

Evaluation is All You Need: Strategic Overclaiming of LLM Reasoning Capabilities Through Evaluation Design cites this paper.

Evaluation is All You Need: Strategic Overclaiming of LLM Reasoning Capabilities Through Evaluation Design Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:40:21.212589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:40:21.212589Z digest=sha256:cb74ce3d69e4642353366de9f0b5f9a3a51077b794d42d079601a8529138fecd

Observation fa67507f-6f74-448c-a283-ebf292b38dd5 · inbound

Com$^2$: A Causal-Guided Benchmark for Exploring Complex Commonsense Reasoning in Large Language Models cites this paper.

Com$^2$: A Causal-Guided Benchmark for Exploring Complex Commonsense Reasoning in Large Language Models Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:46:43.398130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:46:43.398130Z digest=sha256:cdf784b36a722b8fe56c7f7475907af8a589aea676024c75cf37bd6dba4e1037

Observation a4cc418c-f9fe-4f52-9f47-9608d2518715 · inbound

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism cites this paper.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:19.582648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:19.582648Z digest=sha256:86921c210e57b50cd33ea41018a6202933f3708785b310c77f9584d3560640c5

Observation 33c4bf2c-0c93-4766-a506-f9a81f20da20 · inbound

Ctrl-Z Sampling: Scaling Diffusion Sampling with Controlled Random Zigzag Explorations cites this paper.

Ctrl-Z Sampling: Scaling Diffusion Sampling with Controlled Random Zigzag Explorations Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T22:59:12.842671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:59:12.842671Z digest=sha256:6bb4e6478fb539311af8085735cddf40fa386892772c5f4802ea3f74deddf8ed

Observation a1bcfbea-21e6-46dd-96b8-9aa36ae0c15b · inbound

Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs cites this paper.

Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T22:09:46.543585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:09:46.543585Z digest=sha256:a76a481f4b06a69447bb34dd0389ade204a705436ece71f85bd9d277ab54f925

Observation ea005420-8fac-45fb-be45-2bb5acf2547f · inbound

From Trial-and-Error to Improvement: A Systematic Analysis of LLM Exploration Mechanisms in RLVR cites this paper.

From Trial-and-Error to Improvement: A Systematic Analysis of LLM Exploration Mechanisms in RLVR Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T22:03:28.144048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:03:28.144048Z digest=sha256:669a7c93390e0b2997ba25f3d74fb428dc46bc39856cd1579d8d82dbd5db0987

Observation 0a8439d8-5404-4d31-926f-f181fcc433c4 · inbound

Why Does Reasoning Length Converge? Unveiling the Underfitting-Overfitting Trade-off in Chain-of-Thought cites this paper.

Why Does Reasoning Length Converge? Unveiling the Underfitting-Overfitting Trade-off in Chain-of-Thought Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:31:37.800280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:31:37.800280Z digest=sha256:eb79165956e8761bad511bb9ae18eb4a8b0ff584def0f44d2eefab7cbc25710e

Observation 608e598d-cbc1-4caa-a836-3e3fc350f3af · inbound

Sticker-TTS: Learn to Utilize Historical Experience with a Sticker-driven Test-Time Scaling Framework cites this paper.

Sticker-TTS: Learn to Utilize Historical Experience with a Sticker-driven Test-Time Scaling Framework Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T05:45:02.737120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:45:02.737120Z digest=sha256:1dca60f983c664343a0416f5f8dbbbf758aaeafeb0687a240bf5617bbb71c38c

Observation 84507aa2-47d9-400c-a03c-d6af0fc57d8d · inbound

Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts cites this paper.

Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:42:38.635096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T13:42:07.883909Z digest=sha256:c15147ac38b6906ad21ee6b0086528e8a6496a1a40d6d4abe37b161bd717fdd4

Observation 98ab1b38-b39d-4c71-bf79-bb57fba73fd2 · inbound

When control meets large language models: From words to dynamics cites this paper.

When control meets large language models: From words to dynamics Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:54:13.213179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T14:52:44.632671Z digest=sha256:022626096ef92fcbc3a712dd994dadd7126692c2dfe059c6032ba33a52d7a5aa

Observation 537f8259-853b-49fa-bb55-17839a0768ab · inbound

PARM: Pipeline-Adapted Reward Model cites this paper.

PARM: Pipeline-Adapted Reward Model Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.510004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T05:15:26.015817Z digest=sha256:2bcb636902e4be98f913b2837206031ec663dbace603d9882163fbab2df7420c

Observation 3bf1b1a0-94ac-456c-b1eb-0ba4e401864d · inbound

A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability cites this paper.

A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:26:27.522360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T03:03:05.715652Z digest=sha256:7c68423eaac358ea80bd55272bb60312d020774d92bf7eb1b2a5278fd389e997

Observation 0d972d20-b559-42bb-839a-8127afaf352d · inbound

Density Field State Space Models: 1-Bit Distillation, Efficient Inference, and Knowledge Organization in Mamba-2 cites this paper.

Density Field State Space Models: 1-Bit Distillation, Efficient Inference, and Knowledge Organization in Mamba-2 Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:05:37.131006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-01T08:55:37.926334Z digest=sha256:0affd1fa032c0dfe7489c4e39997cfd68f78071784a0f48762aecb3d84a0ccf4

Observation cc708544-57ba-42b6-ab7d-8b6a0db9d35b · inbound

MODE-RAG: Manifold Outlier Diagnosis and Energy-based Retrieval-Augmented Generation Evaluation cites this paper.

MODE-RAG: Manifold Outlier Diagnosis and Energy-based Retrieval-Augmented Generation Evaluation Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:08:56.384800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T01:35:46.913390Z digest=sha256:45d8260f8c8d83dc20625e9d1fb6b71deee5b1319f998317bac3b89986202f68

Observation 7eb9e57f-a20d-44e6-8a98-f55fef3c7839 · inbound

Beyond Fixed Budgets: Characterizing the Inelasticity and Limitations of Tree-of-Thought Reasoning Strategies cites this paper.

Beyond Fixed Budgets: Characterizing the Inelasticity and Limitations of Tree-of-Thought Reasoning Strategies Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:44:59.789825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T18:38:48.512777Z digest=sha256:7ab7e3f9605cc9f7c302d08bdbb98f1adbfa0ba43e700a99430539d4097051f2

Observation 1ecd146f-269f-4c5c-9404-be0d44e64969 · inbound

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning cites this paper.

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 285

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:59:44.800993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T09:19:50.623741Z digest=sha256:76dee599cfb9e1f853df251d9380db08725d4ea88b4eb66d30f0f10527de1f11

Observation e4709ce6-424a-4ecb-92f6-efdc6d7ccd9d · inbound

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning cites this paper.

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 284

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T18:55:59.426665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T01:18:19.195007Z digest=sha256:654c52800edf6568150eeff0020d195a1dbaa778473a023792e6e184cfbf4d34

Observation 40872ebf-2c4b-4d17-ad29-4a171c8d48d5 · inbound

TreeThink: A Modular Tree Search Library for Mathematical Reasoning with LLMs cites this paper.

TreeThink: A Modular Tree Search Library for Mathematical Reasoning with LLMs Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 123

Resolution
unresolved
no resolver link, observed 2026-07-14T05:57:23.399019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T05:57:23.399019Z digest=sha256:0de55bca2195bcfb4e536b7b82e5da8217eea5f716c03119b28a2d8dcad07850