Pith. sign in

Paper Citation Record · LEDGER

Reinforcement Speculative Decoding for Fast Ranking

As of 8 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 3 inbound Pith citation observations for arXiv:2505.20316.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20316 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:55:49.233352Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T04:27:07.923007Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T21:31:51.579040Z

Reference resolution

48 of 48 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cd7e7623-affe-400e-af19-a87abb5b206a · outbound

This paper cites Learning to rank using gradient descent.

Reinforcement Speculative Decoding for Fast Ranking Learning to rank using gradient descent

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:54.309735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:44.807473Z digest=sha256:eeaaf3ad5874aa95a7569c2e18f5a13e3f5c79435b1cd09da07af9faede05d6b

Observation ab55c311-ad0f-4d2c-b5fb-1c1d418b5ce9 · outbound

This paper cites Medusa: Simple llm inference acceleration framework with multiple decoding heads.

Reinforcement Speculative Decoding for Fast Ranking Medusa: Simple llm inference acceleration framework with multiple decoding heads

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:53.981128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:44.877478Z digest=sha256:6294228e0511633f6f1d9df49b6df62e9549fc9b00979ccad051b3c0b9b6c3c9

Observation 951aa2c4-3b1b-48ab-b8b1-b65e47ceea8b · outbound

This paper cites Accelerating Large Language Model Decoding with Speculative Sampling.

Reinforcement Speculative Decoding for Fast Ranking Accelerating Large Language Model Decoding with Speculative Sampling

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:45.008851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:45.008851Z digest=sha256:115371bcb45b8eb82f605f2ec9db5763f9b0c121f07e22db52a6f02e6fc09d0a

Observation 5a260b83-81ef-4d10-b341-a15956c9039a · outbound

This paper cites Sequoia: Scalable, Robust, and Hardware-aware Speculative Decoding.

Reinforcement Speculative Decoding for Fast Ranking Sequoia: Scalable, Robust, and Hardware-aware Speculative Decoding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:45.134263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:45.134263Z digest=sha256:c4defa32f3de77fee8519b1ce3960f2c5bd3e8ec741633fa4a55d11f6f84b4ec

Observation 2cb6909c-ba71-4577-ab3a-5da3a457fc6e · outbound

This paper cites Cascade speculative drafting for even faster llm inference.

Reinforcement Speculative Decoding for Fast Ranking Cascade speculative drafting for even faster llm inference

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:53.698037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:45.203410Z digest=sha256:ad62f4457cf4cf330b1454cb2cc93d19f69605deb9cc87d917e7b3d73a35185c

Observation eda251c9-b5ff-435b-a963-e0037c362d03 · outbound

This paper cites Glide with a cape: a low-hassle method to accelerate speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Glide with a cape: a low-hassle method to accelerate speculative decoding

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:53.440746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:45.309951Z digest=sha256:c7244907e13d898945b3eb74b5c14f87f61f277bf711fb29e34fd478c9033ed5

Observation cbc96cdc-c33b-471a-a449-d53158fdaa20 · outbound

This paper cites Quasi- metric learning for bilateral person-job fit.

Reinforcement Speculative Decoding for Fast Ranking Quasi- metric learning for bilateral person-job fit

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:53.165517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:45.427835Z digest=sha256:257bc635a537296e446f754d1fdab70ad1b5d4b3adbd829d9095b17834a7a52f

Observation b3089c53-0cf5-492e-abc7-86b4c6618674 · outbound

This paper cites Enhancing job recommendation through llm-based generative adversarial networks.

Reinforcement Speculative Decoding for Fast Ranking Enhancing job recommendation through llm-based generative adversarial networks

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:45.509766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:45.509766Z digest=sha256:b74d00c00f08e5f68e3b749dfa6aec521fed653d10f7ec300e0d87a583f18410

Observation 962a438b-8c4e-4740-a585-2a9b5cae7b62 · outbound

This paper cites Active large language model-based knowledge distillation for session-based recommendation.

Reinforcement Speculative Decoding for Fast Ranking Active large language model-based knowledge distillation for session-based recommendation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:52.827209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:45.595587Z digest=sha256:58eb3dd77e196f5cb46aa9e3309b23707858f6de5fe90cf8cccbf4f8b9b8ffe7

Observation 89c8e272-d3ee-4f42-a0b5-8714d252db41 · outbound

This paper cites Break the sequential dependency of llm inference using lookahead decoding.

Reinforcement Speculative Decoding for Fast Ranking Break the sequential dependency of llm inference using lookahead decoding

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:52.580518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:45.690591Z digest=sha256:484f285dbd50c5b2fe1929985d6937fbe78bbfdcb2b9d142496417c2d5e925b6

Observation 1d7f1dd0-0f61-4d9a-b4ea-c07bfa6f41e0 · outbound

This paper cites The Llama 3 Herd of Models.

Reinforcement Speculative Decoding for Fast Ranking The Llama 3 Herd of Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:45.758065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:45.758065Z digest=sha256:ab53be0984b7bdbf85ff1cdcd7d43b5892d0f92b3a6166f9c642672ed4c6d684

Observation f9821ace-78b5-4ccb-ba83-2d9a3b682592 · outbound

This paper cites Rest: Retrieval-based speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Rest: Retrieval-based speculative decoding

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:52.343550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:45.829002Z digest=sha256:bad3ffc7f1788780cdfb954237629bdc7df03619e6316030751d019eaad1d789

Observation df0725ad-3f72-4ec8-9388-84057782f793 · outbound

This paper cites SPEED: Speculative Pipelined Execution for Efficient Decoding.

Reinforcement Speculative Decoding for Fast Ranking SPEED: Speculative Pipelined Execution for Efficient Decoding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:45.915619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:45.915619Z digest=sha256:176cb1a73e8f24e8e74c13e7975bec1f6a036bb33836742150d5b52d0a3df295

Observation a8fc6330-5401-4bd8-bdb5-520adafa04fd · outbound

This paper cites Large language models are zero-shot rankers for recommender systems.

Reinforcement Speculative Decoding for Fast Ranking Large language models are zero-shot rankers for recommender systems

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:52.030308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:46.027434Z digest=sha256:464df8e5461221288cc5754d3e6f8d83a021e919c8d4ffbd0920cdf90253968d

Observation 3b362025-19e4-4368-8fd2-366addca62b1 · outbound

This paper cites Neural input search for large scale recommendation models.

Reinforcement Speculative Decoding for Fast Ranking Neural input search for large scale recommendation models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:46.110093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:46.110093Z digest=sha256:e043273af1d6fbc6e35b6849a2756d51654db579f94a86b68754d0cb67523444

Observation 029d947d-c457-4564-b415-5c9a74fb6ac2 · outbound

This paper cites Speculative decoding with big little decoder.

Reinforcement Speculative Decoding for Fast Ranking Speculative decoding with big little decoder

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:51.858643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:46.201097Z digest=sha256:0274ff7e0f46feb75a14041995899b1253d3c6f6c0c2e2be7aa4f89e03669092

Observation fb07fa24-a7b0-4a81-99f8-db0e151e757a · outbound

This paper cites Ancestral gumbel-top-k sampling for sampling without replacement.

Reinforcement Speculative Decoding for Fast Ranking Ancestral gumbel-top-k sampling for sampling without replacement

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:51.645613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:46.305009Z digest=sha256:734f9109f129a9a65161562a3642b911320a84a95cd8b210058765d2d2477010

Observation ac585f47-9ac5-4718-9f9b-35884071ad15 · outbound

This paper cites Fast inference from transformers via speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Fast inference from transformers via speculative decoding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:46.374306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:46.374306Z digest=sha256:22560cb8aadcf8786a53c4c38a7e8355bcb0f68733e711bcb04721db067ed9f3

Observation d596bb36-bcfc-407d-b4c4-dcb0513c0f27 · outbound

This paper cites Eagle: speculative sampling requires rethinking feature uncertainty.

Reinforcement Speculative Decoding for Fast Ranking Eagle: speculative sampling requires rethinking feature uncertainty

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:51.414693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:46.467020Z digest=sha256:2f7657960cc5960466e0c90cf962590fd9b0d695091fea8db4c969884e0c4b09

Observation 1184db20-b773-42a8-b66a-04d74ab532d8 · outbound

This paper cites Generalized ambiguity decomposition for ranking ensemble learning.

Reinforcement Speculative Decoding for Fast Ranking Generalized ambiguity decomposition for ranking ensemble learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:51.197273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:46.557451Z digest=sha256:3db4dc4759370d0337b2e7e7504e95bb4375ec416d76d90d3af9049bbc6302be

Observation b8570d3b-64b3-491e-a9f2-9008463571c2 · outbound

This paper cites Online speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Online speculative decoding

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:51.105067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:46.700635Z digest=sha256:2dd467eb2cb5945731fd96f64d2e31c47b4a17c5a78b39aef205bc00bd638e73

Observation def8ea7e-72a2-4bd2-90f8-5d6597de3d3b · outbound

This paper cites Ranked list truncation for large language model-based re-ranking.

Reinforcement Speculative Decoding for Fast Ranking Ranked list truncation for large language model-based re-ranking

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:51.010746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:46.802330Z digest=sha256:f44854487b376531bb4df27dbdeb4c291151861405de9c65d257a1ca424d6236

Observation 6f39294c-0397-4870-bea6-5bd41c1dca19 · outbound

This paper cites Specinfer: Accelerating large language model serving with tree-based speculative inference and verification.

Reinforcement Speculative Decoding for Fast Ranking Specinfer: Accelerating large language model serving with tree-based speculative inference and verification

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:46.907373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:46.907373Z digest=sha256:4a19bbcbac017e17149314c438ee991061f5e1581b58eeafabe1a7ab98f37895

Observation 1232ac43-92d0-4be4-bdf3-4324e91ad33f · outbound

This paper cites PaSS: Parallel Speculative Sampling.

Reinforcement Speculative Decoding for Fast Ranking PaSS: Parallel Speculative Sampling

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.016746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.016746Z digest=sha256:a00610c9e565c1cc08864d2a4d1ebc1d700d5b8bb1e4ac898a8fd2cc781f08b0

Observation 0e1972d2-bad1-4713-adaf-c3c589e737b4 · outbound

This paper cites Machine learning: a probabilistic perspective.

Reinforcement Speculative Decoding for Fast Ranking Machine learning: a probabilistic perspective

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.125132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.125132Z digest=sha256:573d8827c8cd4ab809043b186a4d432d7baaf0342e0131238ffa89d97ef77213

Observation 1c218bff-cd9f-48df-b791-b8f3db60fdd4 · outbound

This paper cites Ms marco: A human-generated machine reading comprehension dataset.

Reinforcement Speculative Decoding for Fast Ranking Ms marco: A human-generated machine reading comprehension dataset

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.197655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.197655Z digest=sha256:6b0b69d2093b74c9e1e2b8ee6ef686c23ce4fa5529bc32097797b7df1daa333f

Observation 25477b58-b72e-4a56-ab7f-b3a4cbfee646 · outbound

This paper cites Top-Down Partitioning for Efficient List-Wise Ranking.

Reinforcement Speculative Decoding for Fast Ranking Top-Down Partitioning for Efficient List-Wise Ranking

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.281945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.281945Z digest=sha256:5165b44754d35bb719bc6c7498da1e14b6dd9a296d33a10e70094de5bc271102

Observation fab25013-0afa-47e8-b14f-6c6605413863 · outbound

This paper cites RankVicuna: Zero-Shot Listwise Document Reranking with Open-Source Large Language Models.

Reinforcement Speculative Decoding for Fast Ranking RankVicuna: Zero-Shot Listwise Document Reranking with Open-Source Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.416667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.416667Z digest=sha256:9088e426c3b91aa962df991e47b2455c9165dd67953eb24dcab28c96f6e4b05a

Observation 0fa9f67a-44a5-40cf-9f42-f42d2f6d5a56 · outbound

This paper cites RankZephyr: Effective and Robust Zero-Shot Listwise Reranking is a Breeze!.

Reinforcement Speculative Decoding for Fast Ranking RankZephyr: Effective and Robust Zero-Shot Listwise Reranking is a Breeze!

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.543984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.543984Z digest=sha256:8206deaf04cbad864c7cf3d02aadfcae7660ede1583fa74a96bde3be785fcbc0

Observation 38eb7458-6e76-4f52-832a-668851e71c9a · outbound

This paper cites First: Faster improved listwise reranking with single token decoding.

Reinforcement Speculative Decoding for Fast Ranking First: Faster improved listwise reranking with single token decoding

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.867711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:47.634430Z digest=sha256:d5a78b84806c2e6a0453dc5d517b6f484561c09e48e278181c520e8424ea8483

Observation e0ca8bed-2e0d-4556-94f1-52948ba962f0 · outbound

This paper cites Accelerating transformer inference for translation via parallel decoding.

Reinforcement Speculative Decoding for Fast Ranking Accelerating transformer inference for translation via parallel decoding

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.719645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:47.742704Z digest=sha256:63dfcb619a41baee819b65cd80e19850ec3762303ddd0b8a4f9ba1397e23a86d

Observation c0831dec-029e-4720-a3f0-cd1b6d8bad79 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Reinforcement Speculative Decoding for Fast Ranking DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:47.849400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:47.849400Z digest=sha256:4e1974065dd35d7266a823e38ca5bcdc3f488a0fe786579522eca4cc8037fe19

Observation 42cb5e4f-6e07-47cf-aa8b-216da521f1e6 · outbound

This paper cites Accelerating llm inference with staged speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Accelerating llm inference with staged speculative decoding

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.581792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:47.938052Z digest=sha256:a445a48496e071e03e76b1db2a7f9c2ddea7d40a11fee30d058a194d26fff9ba

Observation 2ed1d937-2e62-4e1d-b113-c74cc172ebed · outbound

This paper cites Blockwise parallel decoding for deep autoregressive models.

Reinforcement Speculative Decoding for Fast Ranking Blockwise parallel decoding for deep autoregressive models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:48.038119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:48.038119Z digest=sha256:b5d0cdbd93c3fc6eeb4c093c16ebecd2f93a3c3107110d515637f3209888fb06

Observation e324952c-13fd-4ee5-b4b4-9cb1bb572cc4 · outbound

This paper cites Spectr: Fast speculative decoding via optimal transport.

Reinforcement Speculative Decoding for Fast Ranking Spectr: Fast speculative decoding via optimal transport

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.454616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:48.131007Z digest=sha256:c42fb676108e54973b807a8eeeca94a1c23dcafcde6fece51c7719dfc0d1ae92

Observation 5543b4e6-97b7-44e7-8170-3a037c85c517 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Reinforcement Speculative Decoding for Fast Ranking LLaMA: Open and Efficient Foundation Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:48.202843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:48.202843Z digest=sha256:d28295acddbc14a4a4dbc29dfcd2f3fbaf697c7fd2112ad663148d2182850476

Observation 00b9c292-7405-4106-9bae-a2ee1930512e · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Reinforcement Speculative Decoding for Fast Ranking Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:48.302963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:48.302963Z digest=sha256:d4217e506a57868d11e6469c91556c9dd9ec58f8f5f902bc311c279488828f99

Observation f4907a0b-19fb-42fd-ad14-e7c5143efc21 · outbound

This paper cites Re2llm: Reflective reinforcement large language model for session-based recommendation.

Reinforcement Speculative Decoding for Fast Ranking Re2llm: Reflective reinforcement large language model for session-based recommendation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.328598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:48.394670Z digest=sha256:1e3b2683a6162a9b06604ef377ec4f6a42bd1bdd6a093ebfa25f96c55cbeff72

Observation 982ed65e-76c0-47ae-823b-8ac58dcdb0ba · outbound

This paper cites A survey on large language models for recommendation.

Reinforcement Speculative Decoding for Fast Ranking A survey on large language models for recommendation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:48.475112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:48.475112Z digest=sha256:cb49c657e4038e194215540a572e70de70eb31d298e0afffc9f99b07a748225a

Observation 2d3dab59-36fc-4ffc-89d2-527ca833f50d · outbound

This paper cites Speculative decoding: Exploiting speculative execution for accelerating seq2seq generation.

Reinforcement Speculative Decoding for Fast Ranking Speculative decoding: Exploiting speculative execution for accelerating seq2seq generation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.195461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:48.534931Z digest=sha256:94b6db8dffe81438ae735bcf25211d6312ca4f28e1a934cb073db1db47cb4549

Observation d903e4ea-3fee-46b8-851d-eaddab119383 · outbound

This paper cites Unlocking efficiency in large language model inference: A comprehensive survey of speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Unlocking efficiency in large language model inference: A comprehensive survey of speculative decoding

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:50.041078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:48.619763Z digest=sha256:919cd56a0cc4021de36c759546aaf7fc9c2e68bb3c6e9420c5a7c296afc4d903

Observation b00634c5-46fa-45d1-a089-0f120fd5358a · outbound

This paper cites Qwen2.5 Technical Report.

Reinforcement Speculative Decoding for Fast Ranking Qwen2.5 Technical Report

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:48.710769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:48.710769Z digest=sha256:b994d63e720388f0f9e8d8f2bf547e371136d6000f0e2fbd181dc568c2639f62

Observation 3fc18fa0-9614-42a1-bc6c-c6d1c1c68ba5 · outbound

This paper cites Multi-Candidate Speculative Decoding.

Reinforcement Speculative Decoding for Fast Ranking Multi-Candidate Speculative Decoding

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:48.798319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:48.798319Z digest=sha256:ea335843373364fb47e1b6495d6535c2e12422ec421bfdfbb6727db14996dd76

Observation d4381522-e859-41c0-8657-75ba81c155eb · outbound

This paper cites Predictive pipelined decoding: A compute-latency trade-off for exact llm decoding.Transactions on Machine Learning Research.

Reinforcement Speculative Decoding for Fast Ranking Predictive pipelined decoding: A compute-latency trade-off for exact llm decoding.Transactions on Machine Learning Research

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:49.919724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:48.853625Z digest=sha256:2cb3063aa867ac55814cecedd045257e0063e479301e17c95fb2416d89a180f4

Observation c469d7b3-9ec3-40ef-98dd-a34d89f5821d · outbound

This paper cites Draft& verify: Lossless large language model acceleration via self-speculative decoding.

Reinforcement Speculative Decoding for Fast Ranking Draft& verify: Lossless large language model acceleration via self-speculative decoding

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:49.807720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:48.949300Z digest=sha256:90d9b67649937f4c616f66a11553f8fb6ef6ad7f3d6b4a87a8b9a8e2478dbed2

Observation f2f37503-db68-4a1d-8215-f8a7edfcb869 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

Reinforcement Speculative Decoding for Fast Ranking OPT: Open Pre-trained Transformer Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:49.016380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:49.016380Z digest=sha256:42eec740b2ef8a5ff741d2140190824067c8c041201ca26de45149d438b6a6ac

Observation fb204efc-0b25-4d70-ae16-6aa166dcc0d1 · outbound

This paper cites Distillspec: Improving speculative decoding via knowledge distillation.

Reinforcement Speculative Decoding for Fast Ranking Distillspec: Improving speculative decoding via knowledge distillation

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:55:49.683573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:55:49.138610Z digest=sha256:192e87cc1292a2f8647a42ced4bef635c342d7a3c0677d9fcdaf825b0334d562

Observation a9040345-ff0a-4f93-975c-c3c4c446655a · outbound

This paper cites Large language models for information retrieval: A survey.

Reinforcement Speculative Decoding for Fast Ranking Large language models for information retrieval: A survey

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:49.233352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:55:49.233352Z digest=sha256:4bb4ece5874a0637d25448367218920cf95ec03259303f17366fb325e2f2167b

Pith citing papers

Observation 062f74ff-c5db-4046-844d-0f849a74c9dc · inbound

Mirroring Users: Towards Building Preference-aligned User Simulator with User Feedback in Recommendation cites this paper.

Mirroring Users: Towards Building Preference-aligned User Simulator with User Feedback in Recommendation Reinforcement Speculative Decoding for Fast Ranking

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-18T21:31:51.582307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T21:28:29.868292Z digest=sha256:c4ec92ef94fd634eda6c2f85ab01677ea577b8d640f766083a6584285a247e5f

Observation d650411c-fd95-43c3-9f9a-5d520b0ad975 · inbound

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs cites this paper.

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs Reinforcement Speculative Decoding for Fast Ranking

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T01:33:51.671261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:33:51.671261Z digest=sha256:5cdfe74b46be0ca4d230cbcbdf56c01ba496827badb3b66e2f0725af6b368090

Observation 157a78b3-e3a3-4441-80ae-23fa8c7516c5 · inbound

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs cites this paper.

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs Reinforcement Speculative Decoding for Fast Ranking

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T04:27:07.923007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:27:07.923007Z digest=sha256:3ccf830ba149cae43dfa84b1bd1bf9285164f337505dc45e8da3dc4f390a54bf