Pith. sign in

Paper Citation Record · LEDGER

Recurrent Drafter for Fast Speculative Decoding in Large Language Models

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2403.09919.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.09919 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:58:58.977065Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 584fac1b-deec-4c13-b56a-37cff770aa98 · inbound

SnapKV: LLM Knows What You are Looking for Before Generation cites this paper.

SnapKV: LLM Knows What You are Looking for Before Generation Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-13T12:57:43.165099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T12:57:43.036674Z digest=sha256:fabe21aef491fb7f4e007c05588c2214022bba5bdf6b73210db64961c0aaf8e1

Observation 2f37e260-1acc-436e-a116-bca6163f732b · inbound

FastDraft: How to Train Your Draft cites this paper.

FastDraft: How to Train Your Draft Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T19:05:26.233764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T19:05:26.233764Z digest=sha256:977f7821745059b7f5ac1ed21c73f887d88c04a1f6bf967640db8dc4329a1935

Observation ace799fd-7be1-4cd7-b1d2-00bc560b6dae · inbound

Judge Decoding: Faster Speculative Sampling Requires Going Beyond Model Alignment cites this paper.

Judge Decoding: Faster Speculative Sampling Requires Going Beyond Model Alignment Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-09T20:43:28.528369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:43:28.528369Z digest=sha256:88a493d133f422c657c68b9c26b76910093fe46a046c4ece0d92e0d4bda2cfad

Observation d3ea340f-0f16-41e1-8fbf-9b2bc3e13228 · inbound

AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset cites this paper.

AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T10:58:58.977065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:58:58.977065Z digest=sha256:93d06274956a4a0f7e5b97294ce0b27072c0f57ad3b0c469444e20c1d152655c

Observation 948327e4-2878-4331-937f-6c0f9abda58f · inbound

Automatic Task Detection and Heterogeneous LLM Speculative Decoding cites this paper.

Automatic Task Detection and Heterogeneous LLM Speculative Decoding Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T21:57:56.230937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:57:56.230937Z digest=sha256:8b0272e2b2e3213315573df1e4852b6f86c109a89097df19252f59d4612c323e

Observation 37362ddd-1f51-42b7-93df-68140225e8db · inbound

Utility-Driven Speculative Decoding for Mixture-of-Experts cites this paper.

Utility-Driven Speculative Decoding for Mixture-of-Experts Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:53.221956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:53.221956Z digest=sha256:5d7cb47d7d53708dc5892bf11c34582dfe5508f902850347dd56a445f31082f7

Observation 960d792a-9245-4d36-8ed9-88a4a24bb524 · inbound

WhisperKit: On-device Real-time ASR with Billion-Scale Transformers cites this paper.

WhisperKit: On-device Real-time ASR with Billion-Scale Transformers Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:29:00.105785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:29:00.105785Z digest=sha256:b823be933766572c08052470a28c9d4548d077abaebd94c7a8f1c71b6ee994a8

Observation fed792ba-de91-4bc2-abbc-1d70904108c2 · inbound

Your LLM Knows the Future: Uncovering Its Multi-Token Prediction Potential cites this paper.

Your LLM Knows the Future: Uncovering Its Multi-Token Prediction Potential Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T17:06:07.543973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:06:07.543973Z digest=sha256:7983f96f064fb4e2aa1788014e9210dcb9fcf996d83a8faabb3e6c645029b542

Observation 9faaa76d-0b1c-4be5-b78f-4da91433aea4 · inbound

SpecASR: Accelerating LLM-based Automatic Speech Recognition via Speculative Decoding cites this paper.

SpecASR: Accelerating LLM-based Automatic Speech Recognition via Speculative Decoding Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T14:42:56.188422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:42:56.188422Z digest=sha256:ea676be545c018c09beae1e5ad5a5c1b70b494c6c0bcc9e8b1302b202a0a2212

Observation e1ffa0f5-d274-4428-801d-dc9c3d3b4b68 · inbound

Speculative Decoding with a Speculative Vocabulary cites this paper.

Speculative Decoding with a Speculative Vocabulary Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T23:27:54.422557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T23:27:54.422557Z digest=sha256:1ce6ef85df7f7036dc86b55ee48eb0e3caa0ac5fe32c6f411041d2b5bb8d71a2

Observation 90208bc0-7082-4ed5-9efb-46549c373407 · inbound

MineDraft: A Framework for Batch Parallel Speculative Decoding cites this paper.

MineDraft: A Framework for Batch Parallel Speculative Decoding Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T21:14:47.037475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T21:14:47.037475Z digest=sha256:7999209612106c716e61323b3440e1657cfe3dec9e671ef31ddeaf01ed02a161

Observation 775420f2-f57c-4e9d-a8d7-2b8ab93ecacf · inbound

SpecBlock: Block-Iterative Speculative Decoding with Dynamic Tree Drafting cites this paper.

SpecBlock: Block-Iterative Speculative Decoding with Dynamic Tree Drafting Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:30:57.295596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T02:27:45.880587Z digest=sha256:ea95fcee0bfc85576d2eef1fe3f6114964104f5bdf49c2cb006a529a27ff915f

Observation 9b3e0ccd-6dcb-4019-b935-863adc69fb00 · inbound

SpecBlock: Block-Iterative Speculative Decoding with Dynamic Tree Drafting cites this paper.

SpecBlock: Block-Iterative Speculative Decoding with Dynamic Tree Drafting Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:51:26.041044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T10:49:25.448580Z digest=sha256:1f54e86be8c58f794839715e2ed03fe497639d3bd563697064b90345c9794014

Observation c9603266-033a-4d2a-b2c5-a251fc5d8665 · inbound

D-PACE: Dynamic Position-Aware Cross-Entropy for Parallel Speculative Drafting cites this paper.

D-PACE: Dynamic Position-Aware Cross-Entropy for Parallel Speculative Drafting Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-20T21:59:06.207309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T21:56:42.264380Z digest=sha256:60b1925c54af3ceb40b4098e0edba364545785b08fc4624d3631910f2ae4f8c6

Observation f8a788a3-1d4f-44ab-bc31-a1d38bc81754 · inbound

TAPS: Target-Aware Prefix Tree Selection for Diffusion-Drafted Speculative Decoding cites this paper.

TAPS: Target-Aware Prefix Tree Selection for Diffusion-Drafted Speculative Decoding Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:22:34.164167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-28T19:13:19.691818Z digest=sha256:7681791ed0a10d2ee5312ecfdbaf688622534aac01b00209e4216fff63f63b23

Observation 0fdb6cb8-3878-4d49-b6c9-785d310cafdd · inbound

Token-Operations-Oriented Inference Optimization Techniques for Large Models cites this paper.

Token-Operations-Oriented Inference Optimization Techniques for Large Models Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 145

Resolution
verified exact
arxiv_id, observed 2026-07-04T05:09:36.572686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T16:15:22.543601Z digest=sha256:8d983df037272c7648ce21e4b3b8549fd7d2b2a57f1de526acd5ab8e0e301035

Observation dcb9e8e1-ea37-48c3-bf08-a7b830dc8b57 · inbound

Token-Operations-Oriented Inference Optimization Techniques for Large Models cites this paper.

Token-Operations-Oriented Inference Optimization Techniques for Large Models Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-02T10:49:15.195678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:49:15.195678Z digest=sha256:1af693f2450c9c77fc161bc4a219ff51652a0e2cd525f229609a474fd7fe02bd

Observation 1eb1b660-0628-489c-b9d7-1b89f67536a5 · inbound

DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation cites this paper.

DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 123

Resolution
unresolved
no resolver link, observed 2026-07-11T08:05:17.460513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T08:05:17.460513Z digest=sha256:bc8dfce24ecfa7035459854a8c459657ac56e1db1355aafe180a3a986010a0e7

Observation 02ecd1c5-79cb-411e-8d02-37f22e333a4f · inbound

Approximate Speculative Decoding cites this paper.

Approximate Speculative Decoding Recurrent Drafter for Fast Speculative Decoding in Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T19:02:12.575698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T19:02:12.575698Z digest=sha256:dceb26bc73b558d9a975824efb0ffda2693c5e15d07ae22c56fe2a3f77c59a86