Pith. sign in

Paper Citation Record · LEDGER

Reward-Guided Speculative Decoding for Efficient LLM Reasoning

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2501.19324.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.19324 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T00:04:00.084214Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bcec02f8-e6ff-473a-8f3b-0cbd1a804409 · inbound

Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models cites this paper.

Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-14T01:29:57.246050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T01:29:56.480020Z digest=sha256:8b8d4dabbefe6b907b34dd8f0db254738908238682244dad86df7b4f90bd0be3

Observation c7ed6251-5e1d-4de9-a306-712159e8ad51 · inbound

Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs cites this paper.

Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:35:13.236480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T05:35:13.118221Z digest=sha256:dcf98df784f0dba85a0f0caf5f1300018de8b472202257bc5c765e511210612f

Observation ea256494-2030-4bbb-ba13-a7d696c27044 · inbound

From Long to Short: LLMs Excel at Trimming Own Reasoning Chains cites this paper.

From Long to Short: LLMs Excel at Trimming Own Reasoning Chains Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T00:04:00.084214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:04:00.084214Z digest=sha256:d8679ac76b9fa0f3053cb3abac668aa6fce4d51323f9a5856a1a4c6833ab9962

Observation f0e449cb-e897-4b6a-9b53-f4b6d1e66884 · inbound

Rethinking Visual Autoregressive Sampling with Information-Grounding Guidance cites this paper.

Rethinking Visual Autoregressive Sampling with Information-Grounding Guidance Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T14:43:13.834755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:43:13.834755Z digest=sha256:e5903575a8116ee2219cee3139d66d5e8ed2b110ed2e43bc270d54c8d9c79828

Observation b28f02d8-d9ba-475a-8e33-0fd3b860cb39 · inbound

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards cites this paper.

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T13:15:44.660507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:15:44.660507Z digest=sha256:ddd6ff9ddc2b6c234d7175ac55680a0b625ed83766a2d633e636ef528ae5f995

Observation 6fcb612f-1c20-47a5-96de-8b4e4f5eda65 · inbound

MixReasoning: Switching Modes to Think cites this paper.

MixReasoning: Switching Modes to Think Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T11:16:36.269089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:16:36.269089Z digest=sha256:5c9f6c079e2256fc273c1417ea8b04429aa2e47780883ada44bfa5f62b4977da

Observation 8a082f96-ef59-4865-b01f-3914f9c8a059 · inbound

Policy-Guided Stepwise Model Routing for Cost-Effective Reasoning cites this paper.

Policy-Guided Stepwise Model Routing for Cost-Effective Reasoning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:01:11.844961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T10:29:23.538806Z digest=sha256:64336ff8d6f4ac7a18d667bfa285453a60e992b6ef1f355e306dd927ed3c75ac

Observation beb4d729-6445-4488-a429-aa9471f96554 · inbound

Post Reasoning: Improving the Performance of Non-Thinking Models at No Cost cites this paper.

Post Reasoning: Improving the Performance of Non-Thinking Models at No Cost Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 115

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:06:09.780452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T10:19:08.451445Z digest=sha256:33fc70c35f79a8b9ffee63086cab5932a9a28565b062c3f60a2b2a75917f5f1c

Observation b1bc2428-7e1f-4d2d-a8f2-886b8997ceac · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:51:30.110828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T03:51:52.375703Z digest=sha256:164dd24ff70013aacbc483302885eacd6b51c35b1487849a73c07524f3f7c679

Observation b9aa2bb9-de64-4d15-9a03-810c6695c4bd · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:15:03.414244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T05:11:32.053440Z digest=sha256:1733feb5faa38f7272c9e272a57b359f48643290e4f6f97218228e300ddbd5de

Observation 6b4b70f5-f254-4976-8fd3-020ac22a6c20 · inbound

Pause and Reflect: Conformal Aggregation for Chain-of-Thought Reasoning cites this paper.

Pause and Reflect: Conformal Aggregation for Chain-of-Thought Reasoning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:23:31.576800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-15T02:23:22.298606Z digest=sha256:c08c8523568eb98a5bb9cfc6952413c38e5c4bf574c03c7ac101b51c4a3f57f5

Observation 6528dc9d-a31c-44d6-937c-719ad32bb994 · inbound

Performance-Driven Policy Optimization for Speculative Decoding with Adaptive Windowing cites this paper.

Performance-Driven Policy Optimization for Speculative Decoding with Adaptive Windowing Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:42:39.750479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T16:40:17.370100Z digest=sha256:5fa5f880e38865e4d47989bb33a2731998340aa16eedc38df502bb5ca6ec9e4a

Observation 8c66803d-ef90-423c-94bd-8205ca26b35d · inbound

ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning cites this paper.

ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:26:28.717883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T10:07:16.700499Z digest=sha256:6b82984078b1c8628fe30d8235aed99793ca89992d168fc4e9683d89e1ac6c5b

Observation 0486a096-6862-4eef-87df-9433fe9830e1 · inbound

KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty cites this paper.

KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-06-27T13:10:55.947824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T13:07:59.509660Z digest=sha256:ad462914353b2093093c309fd124df81bcc93c9df2c051437fd81bbebbcbd75a

Observation d2fb637d-0b15-42bd-9b2f-e1d868ee854d · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 145

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:57:41.512993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:66cc6713f2cf006ce9af0d04bb226aad6fadc8791edfafa52904625d11e5f9fe

Observation c7a64538-8131-4f38-9c2b-110847bc4108 · inbound

Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing cites this paper.

Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:20:06.711514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-25T21:29:40.564852Z digest=sha256:ec4323e373b1ee1b47ccd36d26eaf2686792f6ab27bfcc390c9925039417f225

Observation aac4b144-e9fe-4ae3-a728-0a9b1dbe4f9d · inbound

Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing cites this paper.

Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-01T06:35:29.567228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T06:33:45.824727Z digest=sha256:8c99e8bd42cc6e4ddd4f5ec13628643750d032c0808c6684390c09b63cb172ed

Observation 928758a2-7325-4ee5-a6b3-a370bab1e0a7 · inbound

Before Thinking, Learn to Decide: Proactive Routing for Efficient Visual Reasoning cites this paper.

Before Thinking, Learn to Decide: Proactive Routing for Efficient Visual Reasoning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:34:41.134711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T05:55:19.083517Z digest=sha256:b72930a648b7665508bf0d4e17e470672e9590d8939d462d298c115ce5a413bd

Observation 84169b7c-5669-40d0-9f03-9c4d03678213 · inbound

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making cites this paper.

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T02:42:03.779702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:42:03.779702Z digest=sha256:879351c071e3c959d448fdae6b17f85433d4e1a54f9653f1e97d0b3c02a4366e

Observation 8ba2f2ae-deba-4dbb-a4f7-8e2ef50281cd · inbound

When to Plan: Learning to Select Between Reactive Control and Deliberative Planning cites this paper.

When to Plan: Learning to Select Between Reactive Control and Deliberative Planning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T21:04:09.724503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:04:09.724503Z digest=sha256:2213fabef192570bef03455bb075849533fe8180f2d48aba932bb2778fe83672

Observation 14bc6ab4-22d8-4176-9126-858b46c5137c · inbound

Is Your Model Thinking or Just Stagnating? PUMA: Diagnosing Reasoning Pathology via Phase-Momentum Alignment cites this paper.

Is Your Model Thinking or Just Stagnating? PUMA: Diagnosing Reasoning Pathology via Phase-Momentum Alignment Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T18:49:28.872386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:49:28.872386Z digest=sha256:e0e080ee33717e44458b5ceca627ac6e8a75ef36bc6b2aa98fd0ee68fca1b09c