Pith. sign in

Paper Citation Record · LEDGER

Fast Large Language Model Collaborative Decoding via Speculation

As of 14 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 7 inbound Pith citation observations for arXiv:2502.01662.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.01662 v2

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T19:34:42.923184Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:20:19.183308Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T08:15:32.154273Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy14
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 49871604-9da7-4e22-9b0d-e314f0068ade · outbound

This paper cites GPT-4 Technical Report.

Fast Large Language Model Collaborative Decoding via Speculation GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.718344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.718344Z digest=sha256:aef588f0fd3cbe17c159da5d2e108a0f580722798b921d497b85021fee955673

Observation 7aa4dc56-7a14-49d8-bfd5-489dc2304b41 · outbound

This paper cites Optimized multi-token joint decoding with auxiliary model for llm inference.

Fast Large Language Model Collaborative Decoding via Speculation Optimized multi-token joint decoding with auxiliary model for llm inference

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.573707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.724002Z digest=sha256:240439f15d4aaa4c5950299810296be8abec13c5df452a7fe31d0e411e2e9640

Observation ae8786af-4db5-4dae-ac23-998c72282bbc · outbound

This paper cites Judge decoding: Faster speculative sampling requires going beyond model alignment.

Fast Large Language Model Collaborative Decoding via Speculation Judge decoding: Faster speculative sampling requires going beyond model alignment

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.557785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.729049Z digest=sha256:e75bc68b34ffa3d5cfa34b8414d399a2c82c09c1b95e68094dfc031507b83fc6

Observation 04428c34-ab39-4d2a-ba7b-65186a07b5c2 · outbound

This paper cites Qwen Technical Report.

Fast Large Language Model Collaborative Decoding via Speculation Qwen Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.734335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.734335Z digest=sha256:ab678cd6fa1619ff0505abef6e12f58b6c50b41c1c86fa181fd686e7a9fd2fdc

Observation 8a3a0a72-5f5f-4ef9-b8db-465d01b864f3 · outbound

This paper cites D., Chen, D., and Dao, T.

Fast Large Language Model Collaborative Decoding via Speculation D., Chen, D., and Dao, T

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.541076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.739899Z digest=sha256:911cfee325dfbbc4d3e8fc31d763dfbe6c7c31d1f87f403a1b5bd814e9545646

Observation d5b976c1-93b1-48db-85ef-b1d084624ca6 · outbound

This paper cites FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance.

Fast Large Language Model Collaborative Decoding via Speculation FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.751817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.751817Z digest=sha256:3d35c7e17a0000033a0a2c09163fd2ca389f33f4cd196101e27b8ac3c6fcc493

Observation cf78432b-2df7-41ec-9ecd-33183fcdfe82 · outbound

This paper cites an unresolved cited work.

Fast Large Language Model Collaborative Decoding via Speculation Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.756975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.756975Z digest=sha256:5ae7fc2a939ce1cc0dd7332e2c6b642278ef8d1ae6232cb0cffee21bbbcddd34

Observation a4739842-7586-40bf-8294-ccc7371d628c · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Fast Large Language Model Collaborative Decoding via Speculation Training Verifiers to Solve Math Word Problems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.761775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.761775Z digest=sha256:b8c21ce5904f4445d7e813d7d1ff7a001a680dc2feed77e0f843ff74c91c1d90

Observation eec7d799-9a97-44c5-bc9f-91d2f33b4537 · outbound

This paper cites Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing.

Fast Large Language Model Collaborative Decoding via Speculation Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.766560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.766560Z digest=sha256:3e290a00310111177adf2b8702bebbce84abf1490f953cd808eb9ff963d2b7b6

Observation 252909aa-7469-4e0b-bd55-20a0f166ef01 · outbound

This paper cites The Llama 3 Herd of Models.

Fast Large Language Model Collaborative Decoding via Speculation The Llama 3 Herd of Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.771328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.771328Z digest=sha256:7b71d599ad847dd18a387d447e75644bf2798521b839098ffe32db5d2ff4bce2

Observation 23dd7218-7081-46fd-b675-6f638150924e · outbound

This paper cites Layerskip: Enabling early exit inference and self-speculative decoding.

Fast Large Language Model Collaborative Decoding via Speculation Layerskip: Enabling early exit inference and self-speculative decoding

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.514791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.776079Z digest=sha256:9ee23928f5436ab3d8d34545a8dfc57aa00979d8025d7f433b02e08bc9f4c485

Observation 748291ee-edbe-449d-a385-b5fc5f86673f · outbound

This paper cites Llm-topla: Efficient llm ensemble by maximising diversity.

Fast Large Language Model Collaborative Decoding via Speculation Llm-topla: Efficient llm ensemble by maximising diversity

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.498782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.780629Z digest=sha256:ae82546dc0527ad92854033fd69f18c9aa008ae9cfa53042e72f22e795a25c4c

Observation d40c8953-5ab6-4f82-bfdb-b9a654dced8d · outbound

This paper cites Graph-structured speculative decoding.

Fast Large Language Model Collaborative Decoding via Speculation Graph-structured speculative decoding

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.482971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.785156Z digest=sha256:5cdde8d500bc5f0c245ead1d56200fe2d7c201018d890aa0d9c80c327bc3cb71

Observation a33534c7-7f12-43c4-8123-8d23427725b5 · outbound

This paper cites Measuring massive multitask language understanding.

Fast Large Language Model Collaborative Decoding via Speculation Measuring massive multitask language understanding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.789550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.789550Z digest=sha256:0a8e3d85f0937fa7327b8afd3519b9ba5fa16a42645f49236cadc3319b48e55e

Observation 75fa5e1c-755e-483e-92e0-46360856478e · outbound

This paper cites and Huang, H.

Fast Large Language Model Collaborative Decoding via Speculation and Huang, H

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.454728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.794003Z digest=sha256:2d24aadc685dfc8d8039b26f49b46ac25918a77fb420ecd36fa1d023bf00725f

Observation b6bf4570-b47d-4935-b25b-c2f6001300c1 · outbound

This paper cites an unresolved cited work.

Fast Large Language Model Collaborative Decoding via Speculation Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.798452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.798452Z digest=sha256:e37d73deb54c2276af384964bf420b336bdc56392edb7f2f91e135e915927f4e

Observation fb7fb685-553c-47a5-a86f-386059aa2797 · outbound

This paper cites an unresolved cited work.

Fast Large Language Model Collaborative Decoding via Speculation Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-09T19:34:43.427208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.803010Z digest=sha256:ed0dc04ab418afdf154a553730bca4b70725e79762174e03d6df5dcbc3307748

Observation fb261483-57e6-4297-a5e1-62adc71cd95d · outbound

This paper cites Fast inference from transformers via speculative decoding.

Fast Large Language Model Collaborative Decoding via Speculation Fast inference from transformers via speculative decoding

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.411677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.807337Z digest=sha256:c84c86bf9d9098c8e0fca87ba40f7c54c56256a2411476b86e1df9597fa86d4d

Observation e412d72e-990f-4846-b9d3-f9aedcdc5b8e · outbound

This paper cites Purifying Large Language Models by Ensembling a Small Language Model.

Fast Large Language Model Collaborative Decoding via Speculation Purifying Large Language Models by Ensembling a Small Language Model

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.813056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.813056Z digest=sha256:8cbbeb17328abd0f7e1a75379734d1891ccba762dcc92e94a783e1e5fe2168a4

Observation 0e2fa5d1-4094-48c8-b91f-d876e514fcb6 · outbound

This paper cites L., Holtzman, A., Fried, D., Liang, P., Eisner, J., Hashimoto, T.

Fast Large Language Model Collaborative Decoding via Speculation L., Holtzman, A., Fried, D., Liang, P., Eisner, J., Hashimoto, T

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.394793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.818122Z digest=sha256:5490973bfcb8d8d4399c1456fd9a6d2385b208da7f579e2eeb1071aec5152317

Observation 67c3f56f-e88c-46de-993f-27dfadf9262b · outbound

This paper cites Eagle: Speculative sampling requires rethinking feature uncertainty.

Fast Large Language Model Collaborative Decoding via Speculation Eagle: Speculative sampling requires rethinking feature uncertainty

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.379561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.822769Z digest=sha256:06ee83bb27d2b104540b5b29d839bf2e2d60afeccecf8e439e287026bf21e1de

Observation 4b49196a-229e-40f8-82bd-750fbf7533be · outbound

This paper cites Eagle-2: Faster inference of language models with dynamic draft trees.

Fast Large Language Model Collaborative Decoding via Speculation Eagle-2: Faster inference of language models with dynamic draft trees

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.364261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.827314Z digest=sha256:c3a4c954607c03a72e8c54b21e5b41bdb0722b8e3f3b78add38710a3a58d49e4

Observation 48bfff0f-49b9-468d-981c-56d669518c22 · outbound

This paper cites DeepSeek-V3 Technical Report.

Fast Large Language Model Collaborative Decoding via Speculation DeepSeek-V3 Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.831914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.831914Z digest=sha256:20054a9b8c9f51ac48124265f6a0d2e7e633f250139132276c2943b6ffef0155

Observation 02f97edf-a985-47ca-a27f-00e711e5f0be · outbound

This paper cites Merge, Ensemble, and Cooperate! A Survey on Collaborative Strategies in the Era of Large Language Models.

Fast Large Language Model Collaborative Decoding via Speculation Merge, Ensemble, and Cooperate! A Survey on Collaborative Strategies in the Era of Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.836845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.836845Z digest=sha256:2caa8ea542ea4c14b8cff58c9d6774a0ca2e6fe4e9be15ea96a5fe5c9cc298e2

Observation d2b8a880-4c2a-4ee9-a89f-fe0377bb164d · outbound

This paper cites Routing to the Expert: Efficient Reward-guided Ensemble of Large Language Models.

Fast Large Language Model Collaborative Decoding via Speculation Routing to the Expert: Efficient Reward-guided Ensemble of Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.841795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.841795Z digest=sha256:ceeaa3666a832c52ecce762cc4ad1d27a0caa26da459f5c4d50f11dade3bc8af

Observation 9bdd421b-f85f-437f-a567-40302a9d35e4 · outbound

This paper cites Blending Is All You Need: Cheaper, Better Alternative to Trillion-Parameters LLM.

Fast Large Language Model Collaborative Decoding via Speculation Blending Is All You Need: Cheaper, Better Alternative to Trillion-Parameters LLM

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.846637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.846637Z digest=sha256:f17e7eae5313bdc14d972f27eae3b75df9c7f9a297cf95facd7f4eb40ab9522b

Observation 72939cfb-f673-4b09-9f62-2d5fa45a280a · outbound

This paper cites an unresolved cited work.

Fast Large Language Model Collaborative Decoding via Speculation Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-09T19:34:43.348336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.851824Z digest=sha256:13c7bd96403f89b02fefdb02b14195ec3b6cc1a57f77689153780d4a620cefa1

Observation 77f8a8d3-87c7-455c-954d-c70c8b7efd55 · outbound

This paper cites Accelerating Large Language Model Decoding with Speculative Sampling.

Fast Large Language Model Collaborative Decoding via Speculation Accelerating Large Language Model Decoding with Speculative Sampling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.856186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.856186Z digest=sha256:11636aff88d84ba523dbeb4276b93e683a873ce5dd30be28d20fcae8b10f3aed

Observation 03a5b853-f377-4768-98d3-6931b2cf520e · outbound

This paper cites J., and Manning, C.

Fast Large Language Model Collaborative Decoding via Speculation J., and Manning, C

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.860706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.860706Z digest=sha256:16e882ecf83943154f1abc6fd7c5d8fd9a36eef76d69cf98620d648b474e3822

Observation 27ebf683-fdb2-41a6-ac28-69759eb4949c · outbound

This paper cites Large language model routing with benchmark datasets.

Fast Large Language Model Collaborative Decoding via Speculation Large language model routing with benchmark datasets

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.332303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.865522Z digest=sha256:d653557026b723dfe0a5434783b3e69b2f1334fb4c7463e5e5b700085b233c55

Observation b2c7f695-12f1-424e-8de9-df01264d1df0 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Fast Large Language Model Collaborative Decoding via Speculation Gemini: A Family of Highly Capable Multimodal Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.870039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.870039Z digest=sha256:106c9af2f8adb4735c0ebdfb8ac02260c4f022b07dfbc73806b0650f5cf6708b

Observation 05fe3f58-a276-4ef0-8fb0-82ff86f795d5 · outbound

This paper cites Qwen2.5: A party of foundation models, September 2024.

Fast Large Language Model Collaborative Decoding via Speculation Qwen2.5: A party of foundation models, September 2024

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.874887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.874887Z digest=sha256:57c1908fcc7f0f3d0eb74fdb1455157d26a20c1e0073d6d74a6c8a1089f2606f

Observation 090bae2a-49ad-4bf2-8af7-c14f962c974b · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Fast Large Language Model Collaborative Decoding via Speculation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.879635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.879635Z digest=sha256:1b884319f728c8a43e2f60b30f43daac459962c419e0b9027a31e4d5b2aaece9

Observation a18cd2b1-401e-46a9-a9ed-6ccf3ad43396 · outbound

This paper cites MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation.

Fast Large Language Model Collaborative Decoding via Speculation MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.884194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.884194Z digest=sha256:dcaa8b1b9c924969442c9cc2669a6016818a4bb1ea3d86e3ce8bf3610045b5b3

Observation 2ea745e2-d4ab-4b09-9eec-151ba1c3f0d7 · outbound

This paper cites Generation Meets Verification: Accelerating Large Language Model Inference with Smart Parallel Auto-Correct Decoding.

Fast Large Language Model Collaborative Decoding via Speculation Generation Meets Verification: Accelerating Large Language Model Inference with Smart Parallel Auto-Correct Decoding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.889215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.889215Z digest=sha256:d2020c068f246301cb938f3ffb3fb386258047c1e40acdb6de747d25aa439b95

Observation 1127ae5f-789a-4bfa-bd84-6e73b7bd718e · outbound

This paper cites C., Ziqi, Y., Yucheng, C., and Li, Y.-S.

Fast Large Language Model Collaborative Decoding via Speculation C., Ziqi, Y., Yucheng, C., and Li, Y.-S

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.894227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.894227Z digest=sha256:3779e154d7742569bcade0aaa4f18be99bf8e6533a8c90582bd149f79e399569

Observation eb5975f4-8907-4908-ab98-6f337af84a9f · outbound

This paper cites Speculative contrastive decoding.

Fast Large Language Model Collaborative Decoding via Speculation Speculative contrastive decoding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.899070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.899070Z digest=sha256:0f8e22b3f1bb635483ea7bb397fa9259c53f1d1c882b3905e6cc22a23b7ae22e

Observation b9de50db-3336-40d3-915a-dcebf95d5daf · outbound

This paper cites Draft&verify: Lossless large language model acceleration via self-speculative decoding.

Fast Large Language Model Collaborative Decoding via Speculation Draft&verify: Lossless large language model acceleration via self-speculative decoding

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.305285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.903885Z digest=sha256:64ac9663c14518e7d6cb59324882d1b565fceca88989750d5493ef5f07cc129b

Observation fd20c50e-dc9e-4d74-bfff-c787c4675911 · outbound

This paper cites V., Mihaylov, T., Ott, M., Shleifer, S., Shuster, K., Simig, D., Koura, P.

Fast Large Language Model Collaborative Decoding via Speculation V., Mihaylov, T., Ott, M., Shleifer, S., Shuster, K., Simig, D., Koura, P

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.908917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.908917Z digest=sha256:833f065f6e767cddb6a7f6ab2c5e67469524f1f2b01e3ef9a9c8e98d1a18de79

Observation 30691e11-0381-4fe2-9c08-9a63eecc2e8a · outbound

This paper cites P., Zhang, H., Gonzalez, J.

Fast Large Language Model Collaborative Decoding via Speculation P., Zhang, H., Gonzalez, J

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.913554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.913554Z digest=sha256:d60f7cdb982a00c6a9d4ad07aea9f215da14d75c0c7dc25c7ae09f45ea7e4d89

Observation 11d3e7c9-1227-4859-ae1a-6560d4efc7e6 · outbound

This paper cites S., Menon, A., Rostamizadeh, A., Kumar, S., Kagy, J.-F., and Agarwal, R.

Fast Large Language Model Collaborative Decoding via Speculation S., Menon, A., Rostamizadeh, A., Kumar, S., Kagy, J.-F., and Agarwal, R

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.267902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.918529Z digest=sha256:bc478a50e7118ca44b845791824c32d95e028931b6cf3f8cfdc3c5e1aeb49e44

Observation c9fcbd5e-9ba1-411f-8df0-3a967ae10fac · outbound

This paper cites write newline.

Fast Large Language Model Collaborative Decoding via Speculation write newline

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.923184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.923184Z digest=sha256:3adf5eef0a490ea8f0092393fea12af39d9a5a455b2a558821b962824ba1448d

Pith citing papers

Observation c7344fe2-73aa-4952-8948-c3ad8ce35ac5 · inbound

PBI-Attack: Prior-Guided Bimodal Interactive Black-Box Jailbreak Attack for Toxicity Maximization cites this paper.

PBI-Attack: Prior-Guided Bimodal Interactive Black-Box Jailbreak Attack for Toxicity Maximization Fast Large Language Model Collaborative Decoding via Speculation

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T20:20:19.183308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:20:19.183308Z digest=sha256:cb5f42024e516b30eaf80151e64c03494781284ae6a6640b398652305576d6c7

Observation 571f5654-e954-4b8c-aeb2-2456940e7102 · inbound

Multimodal Tabular Reasoning with Privileged Structured Information cites this paper.

Multimodal Tabular Reasoning with Privileged Structured Information Fast Large Language Model Collaborative Decoding via Speculation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:53:58.892122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:53:58.892122Z digest=sha256:c587fe9219a0f839969473aeaddd484eca310bb7bcda0413032bf49301356778

Observation 54d16537-58a6-40ee-901d-fa6a88501dbf · inbound

SignAligner: Harmonizing Complementary Pose Modalities for Coherent Sign Language Generation cites this paper.

SignAligner: Harmonizing Complementary Pose Modalities for Coherent Sign Language Generation Fast Large Language Model Collaborative Decoding via Speculation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:11.741454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:11.741454Z digest=sha256:714fb512984bafeb1d739105aa7714811fd56e00125c7081a041c8ee9d79c46b

Observation 684a0bb9-e8f3-4fb3-98b2-7ac8e6982251 · inbound

AnchorSeg: Language Grounded Query Banks for Reasoning Segmentation cites this paper.

AnchorSeg: Language Grounded Query Banks for Reasoning Segmentation Fast Large Language Model Collaborative Decoding via Speculation

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:43:49.444527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-10T05:10:44.608959Z digest=sha256:02cde005db7fda01a5f5150978a51a2ab0fcae8fd855ef8ffff5deb5f41f932e

Observation 7ed78080-faf4-480a-b385-14755324412a · inbound

SpecFed: Accelerating Federated LLM Inference with Speculative Decoding and Compressed Transmission cites this paper.

SpecFed: Accelerating Federated LLM Inference with Speculative Decoding and Compressed Transmission Fast Large Language Model Collaborative Decoding via Speculation

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:31:17.251029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-07T15:12:12.450979Z digest=sha256:f3e9b1114bd92db4a50f9e5a1bc7a15efb45ae0e9408f0eaaa3ab148f7b2b83e

Observation 4c1f544c-6ebb-4f79-ab83-08e5ecbad181 · inbound

Rethinking LLM Ensembling from the Perspective of Mixture Models cites this paper.

Rethinking LLM Ensembling from the Perspective of Mixture Models Fast Large Language Model Collaborative Decoding via Speculation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:15:32.157274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-01T08:07:09.032028Z digest=sha256:8902463ddfa2edc09e0c42852e8a67ccfb4386c38396518ae3bf377858102947

Observation 9f275b1b-19c2-4763-b4dc-173b72f56d06 · inbound

Revisiting Lossy Verification in Speculative Decoding: Mechanisms, Trade-offs, and Failure Modes cites this paper.

Revisiting Lossy Verification in Speculative Decoding: Mechanisms, Trade-offs, and Failure Modes Fast Large Language Model Collaborative Decoding via Speculation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-01T12:07:01.364373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T12:07:01.364373Z digest=sha256:fd52a8cce9a6beff8d0fac6852042d163e505c9cf0a03bc69145dab9f5a8ff36