Pith. sign in

Paper Citation Record · LEDGER

The Synergy of Speculative Decoding and Batching in Serving Large Language Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2310.18813.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.18813 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:21:57.801138Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1d9d6b26-aa50-4313-b1e3-048f81afd2bc · inbound

Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads cites this paper.

Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads The Synergy of Speculative Decoding and Batching in Serving Large Language Models

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-13T10:36:18.056478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T10:36:17.764761Z digest=sha256:1056053d1345216bc8664613371d3a536a889256fb07e0d640bfe50f1b25593b

Observation 9e8243d1-1db2-46a3-9b28-338b160e2010 · inbound

DECA: A Near-Core LLM Decompression Accelerator Grounded on a 3D Roofline Model cites this paper.

DECA: A Near-Core LLM Decompression Accelerator Grounded on a 3D Roofline Model The Synergy of Speculative Decoding and Batching in Serving Large Language Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:57.801138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:57.801138Z digest=sha256:d97192ba80eb7c9984859a5fe0772243a2414173246de0d704e3748348929670

Observation d13dfcad-6739-4a19-afb7-61714e8215ec · inbound

Efficient Speculative Decoding for Llama at Scale: Challenges and Solutions cites this paper.

Efficient Speculative Decoding for Llama at Scale: Challenges and Solutions The Synergy of Speculative Decoding and Batching in Serving Large Language Models

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-05T21:38:54.729462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:38:54.729462Z digest=sha256:cbbfc1d8f57995f1dccaeb4e4b782afb4ae21fa88688a330c452e846c1c4bcda

Observation 32449860-2d8f-494f-b064-adc8a2003669 · inbound

Multi-SPIN: Multi-Access Speculative Inference for Cooperative Token Generation at the Edge cites this paper.

Multi-SPIN: Multi-Access Speculative Inference for Cooperative Token Generation at the Edge The Synergy of Speculative Decoding and Batching in Serving Large Language Models

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T11:06:53.285938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T04:41:10.270844Z digest=sha256:913efa3809973e93fe6642a4bf8a464463a8816acb47c4c9b878534a4983af95

Observation 6c74206b-f3dd-48eb-b5a5-6c38da38c923 · inbound

Sparrow: Sparse Rollout for Stable and Efficient Long-context RL of Large Language Models cites this paper.

Sparrow: Sparse Rollout for Stable and Efficient Long-context RL of Large Language Models The Synergy of Speculative Decoding and Batching in Serving Large Language Models

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:07:26.497068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T19:10:53.882876Z digest=sha256:047350955bd3ea77dd179841160089f62cbb69cbe723bb204e9038aa67ca7a18

Observation 634edaa1-0815-40a4-a020-26ddf9523be5 · inbound

When Is a Draft Accepted? A Theory of Acceptance in Speculative Decoding cites this paper.

When Is a Draft Accepted? A Theory of Acceptance in Speculative Decoding The Synergy of Speculative Decoding and Batching in Serving Large Language Models

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:44:20.993398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-30T07:39:38.006860Z digest=sha256:355e3fae608f7024122d6c0cdd16afe308d2ab56050b1aa920aeeec10df1e989