Pith. sign in

Paper Citation Record · LEDGER

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation

As of 6 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2604.27747.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.27747 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-07T07:42:36.852261Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact14
  • verified fuzzy45
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ecd621d7-e368-4333-8ae5-93293caad541 · outbound

This paper cites arXiv preprint arXiv:2509.03236 , year=.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation arXiv preprint arXiv:2509.03236 , year=

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T10:06:29.872890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:35ac7fb1e83fe67aa6d6845de1a995407a120e05aa26612df975bfaf52a3338c

Observation fe900949-ae80-4c7b-95c3-6a140dfc5d45 · outbound

This paper cites OneRec: Unifying Retrieve and Rank with Generative Recommender and Iterative Preference Alignment.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation OneRec: Unifying Retrieve and Rank with Generative Recommender and Iterative Preference Alignment

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:30:36.120082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:c841820c5d74d2e6026a251690f8619d0713eaac32723a81fcc8755b1a1ac596

Observation 34d86ddc-4448-4692-8736-ebda98ad9261 · outbound

This paper cites Dlcrec: A novel approach for managing diversity in llm-based recommender systems.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Dlcrec: A novel approach for managing diversity in llm-based recommender systems

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.153021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:2b1e0331185c29a179c4c41615ef6d030bd6f1501b46e0eeeedfdf3d85901c53

Observation 05c500de-309d-4449-9a99-1ecb9532bbdf · outbound

This paper cites Knowledge-enhanced con- versational recommendation via transformer-based sequential modeling.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Knowledge-enhanced con- versational recommendation via transformer-based sequential modeling

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.146740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:08288050d8b8e8790eb97ec1ecf0ec2e8cc4e8ff2633c5e642f69b6419e7d990

Observation 81f784fe-558c-4c0d-be42-b7ccf3609845 · outbound

This paper cites Decoding in latent spaces for efficient inference in llm-based recommendation.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Decoding in latent spaces for efficient inference in llm-based recommendation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:29.855506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:828895e5958d2b7f38e6de0b1d0c7e5c8f9b4a08f836b629592a31e103752d12

Observation 4ed9aa09-4465-4552-b0a7-7a6a6e9858b4 · outbound

This paper cites Efficient inference for large language model-based generative recommendation.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Efficient inference for large language model-based generative recommendation

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.269224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:7e930db12cbf2f686e45f2cdee41412f99d35272891e8d67d4363f3b472fd9ed

Observation 66a6b10b-2912-40d9-98bd-1fb57cd933ed · outbound

This paper cites Efficiency unleashed: Inference acceleration for llm-based recommender systems with speculative decoding.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Efficiency unleashed: Inference acceleration for llm-based recommender systems with speculative decoding

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.137408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:256e9eaee446ae6789e6488693161cfbb7d8d632f91b039793ad7f2cbb0335da

Observation 78111f0a-7d0a-47d0-b84e-cc83e83f5b69 · outbound

This paper cites Fast inference from transform- ers via speculative decoding.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Fast inference from transform- ers via speculative decoding

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.255101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:d36891eb6838557885d195d69375c7b354d300f76665adfb8fa7b61484552e71

Observation 52d294ab-ba12-4e7c-9c26-463ec80f03b4 · outbound

This paper cites Speculative Decoding: Exploiting Speculative Execution for Accelerating Seq2seq Generation.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Speculative Decoding: Exploiting Speculative Execution for Accelerating Seq2seq Generation

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T10:06:29.908758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:3da8f233392dd9218e439c58a57703190ccbf7efd2dd15ec69626fc9ae8ae393

Observation 0023b997-b4a0-478d-99b5-92a106d8ed04 · outbound

This paper cites Unlocking efficiency in large language model inference: A comprehensive survey of speculative decoding.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Unlocking efficiency in large language model inference: A comprehensive survey of speculative decoding

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.286173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:873ea4aa2090dccb5119f77dddc75a1bd1f034fb5580d5aca87eb3fa5404c7b1

Observation d04f00d6-ec81-44b4-9f0a-3f9a4c10f530 · outbound

This paper cites Accelerating Large Language Model Decoding with Speculative Sampling.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Accelerating Large Language Model Decoding with Speculative Sampling

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-12T10:06:29.891667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:c8f2534739fdad2c7e8899e63da685f9080ff65d068f92bd1c29e0365bd029fd

Observation ee4d2755-e4e4-451c-a62b-6ad3bb613513 · outbound

This paper cites Specinfer: Accelerating large language model serving with tree-based speculative inference and ver- ification.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Specinfer: Accelerating large language model serving with tree-based speculative inference and ver- ification

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.250871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:b22ae11dec22f7e4fd42c28fe66619383c1e5516f12df48736a17a695358b28d

Observation 20436695-608a-40d5-901a-457577fd7c1c · outbound

This paper cites DistillSpec: Improving Speculative Decoding via Knowledge Distillation.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation DistillSpec: Improving Speculative Decoding via Knowledge Distillation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:29.899708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:0871144e2368209122f093ffadacde2ce5d2bb0157a3f306c310fe57863d3681

Observation 73990dbe-a379-461a-80b9-312fc25d7f16 · outbound

This paper cites Eagle: Speculative sampling requires rethinking feature uncertainty.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Eagle: Speculative sampling requires rethinking feature uncertainty

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.237330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:5057ff644f9105d92b9f7d4b7b6c36db992867bf0d017c3dd4f29350c99bf48a

Observation 5ad440f3-6f73-45d7-8eda-bc97bc154366 · outbound

This paper cites Learning harmonized rep- resentations for speculative sampling.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Learning harmonized rep- resentations for speculative sampling

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.222469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:e061da2182622830e6d7c46c5c3c8994885cb7ebdc4e47394a7c442a4d1d0261

Observation 1aa4c6c1-1440-4e4b-98e0-13489a045b1c · outbound

This paper cites Tokenrec: Learning to tokenize id for llm-based generative recommendations.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Tokenrec: Learning to tokenize id for llm-based generative recommendations

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.205906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:44e1429bf95f5896cfd877a9845e2acba5d8eba1bcc76cac66a514a8654f843b

Observation 546e092e-6081-4435-a44f-5de73ae45449 · outbound

This paper cites Adapting large language models by integrating collaborative semantics for recommendation.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Adapting large language models by integrating collaborative semantics for recommendation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.202444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:098246fadb29e55041642706e15d1ac4955212e044e40d923ffe2b85a2bbbcd6

Observation 6d7b01e3-6c73-4fec-89a3-b6e686d4b8ba · outbound

This paper cites Autoregressive image generation using residual quantization.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Autoregressive image generation using residual quantization

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.172514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:8da1f087c11131f14b58b8aea6952804863f40d9c41a7c9b42ef638502ff128a

Observation 12140ae8-0e99-41cd-8568-95d2d61d1740 · outbound

This paper cites Blockwise parallel decoding for deep autoregressive models.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Blockwise parallel decoding for deep autoregressive models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.272339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:12beff6b16a61193a759d1c25baa48bc21d1250176ab4bf90d4ced8bbdb6e12d

Observation 5806231e-8c2c-4088-b2f6-11c00b27ff9e · outbound

This paper cites Eagle-2: Faster inference of language models with dynamic draft trees.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Eagle-2: Faster inference of language models with dynamic draft trees

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.133117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:5a0f91348ac330a27a86424ed6f76ba6e966a825fdba225d70fbe513e733e88a

Observation 615d321e-d21c-4849-a3d1-bafd50b10521 · outbound

This paper cites Sequoia: Scalable, robust, and hardware-aware speculative decoding.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Sequoia: Scalable, robust, and hardware-aware speculative decoding

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.129079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:8207f0a8a05749efc8f8eac37befb7a3ef59bdbacb756d5a3d7707517574d1ca

Observation ae4474a0-e374-4643-ade2-4a24c20469e2 · outbound

This paper cites Specexec: Massively parallel speculative decoding for interactive llm inference on consumer devices.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Specexec: Massively parallel speculative decoding for interactive llm inference on consumer devices

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.275663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:2040ea73a27222fee2c643ec932ba88e690076b7a2de364b8f7c8e6a38d2521f

Observation 0158217d-1e6b-473b-98f7-ecee20737343 · outbound

This paper cites Break the sequential de- pendency of llm inference using lookahead decoding.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Break the sequential de- pendency of llm inference using lookahead decoding

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.289044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:a0f524e6ac17d53e95c33322c2638bc63a68306f124d670569ca8c0e74d0f38e

Observation 80ee3c1b-e644-4373-a042-75603652d67f · outbound

This paper cites Cllms: Consistency large language models.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Cllms: Consistency large language models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.240744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:d84b0a2bb8f4acb41344d0c907eb730d67c2d0b56504c2bd74169be1ddcbd2cd

Observation 413e5b16-0e87-41d2-92f0-51147bdcbebc · outbound

This paper cites Rest: Retrieval- based speculative decoding.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Rest: Retrieval- based speculative decoding

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.258525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:62b36b13d474d31bb7e953655d2c1dd72f49186af55d2138bf95f2e5da0060ef

Observation 1fb7d90f-2a76-444e-bc95-59cc2b53bad9 · outbound

This paper cites Ouroboros: Generating longer drafts phrase by phrase for faster speculative decoding.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Ouroboros: Generating longer drafts phrase by phrase for faster speculative decoding

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.176731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:01c0ac90ed0f6e50bff99329ee1113b9c3976963675b67ee1aca8f8913e440c0

Observation 73db22a3-fc50-4b9f-940f-d39fd897ef9e · outbound

This paper cites Medusa: Simple llm inference acceleration framework with multiple decoding heads.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Medusa: Simple llm inference acceleration framework with multiple decoding heads

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.219564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:4a127e70ea4d98a69178be4a5eb7eadca0d09932b4f90b37e6add9d58468432b

Observation 3cd53210-084a-42ee-a65e-938242bfa5b3 · outbound

This paper cites Hydra: Sequentially-dependent draft heads for medusa decoding.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Hydra: Sequentially-dependent draft heads for medusa decoding

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.278832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:da5d69b8ac484c70b7a640d3fe7251e91e739d5d85f90d9fa8afd4960fb9c54b

Observation 8a298b37-7ab6-4baf-b354-6a9ca9f15185 · outbound

This paper cites Griffin: Effective token alignment for faster speculative decoding.arXiv preprint arXiv:2502.11018, 2025a.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Griffin: Effective token alignment for faster speculative decoding.arXiv preprint arXiv:2502.11018, 2025a

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:30.198524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:d960f2363adae3c14771e9b9e890c768289e685ab395cae201ecb8c71559ceae

Observation b1032ac3-1c44-4fed-8111-ba6b5f4d0ba9 · outbound

This paper cites Boosting Lossless Speculative Decoding via Feature Sampling and Partial Alignment Distillation.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Boosting Lossless Speculative Decoding via Feature Sampling and Partial Alignment Distillation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:30.243035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:bfca7f0b08e34ec19eb1538a60157bc0b1895811730fb2929d4043cddcbf084a

Observation 5e5a82c0-9f92-4c46-94ec-3a2516ebc520 · outbound

This paper cites CORAL: Learning Consistent Representations across Multi-step Training with Lighter Speculative Drafter.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation CORAL: Learning Consistent Representations across Multi-step Training with Lighter Speculative Drafter

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:29.865081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:e596d3bd455771bb85f7b0a5568537f9c6d5b56f04ac20aedc6787a44d797cd7

Observation b8a2e4bf-96cf-48a7-b633-5bff191870f7 · outbound

This paper cites How speculative can speculative decoding be?.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation How speculative can speculative decoding be?

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.180538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:79a5113803e5854a6100a340f13bce1e89289e68ac92fe132f1cac76b19d636c

Observation fd82ee0f-7577-4b08-b95d-b329e92b1c80 · outbound

This paper cites Heterospec: Leveraging con- textual heterogeneity for efficient speculative decoding.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Heterospec: Leveraging con- textual heterogeneity for efficient speculative decoding

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:29.886601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:4e1a3e4aaf1ffdc305fc1431624dc97d000dac17f5b0890f12b03bf7972528af

Observation 9690de99-d5b6-44e1-9d36-6770af8ec623 · outbound

This paper cites H2o: Heavy-hitter oracle for efficient generative inference of large language models.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation H2o: Heavy-hitter oracle for efficient generative inference of large language models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.195060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:797e0a7b6549368c9a8c00ce167f912563c37268ad9a72c042cda56b74256585

Observation 8633d075-2006-4ee0-8a8e-7dbe81562684 · outbound

This paper cites A bi-step grounding paradigm for large language mod- els in recommendation systems.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation A bi-step grounding paradigm for large language mod- els in recommendation systems

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.186470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:3de8591fb4b406c7b8bdebc9b295b9429992d963fb0bec05b71d92fc0af36fd1

Observation e6ee1658-c903-4837-a11d-d39d042f7526 · outbound

This paper cites Large language models are learnable planners for long-term recommendation.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Large language models are learnable planners for long-term recommendation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.161057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:7b07379a851a3ea5b04ee935379804ebd070b78f4973d9f6b42957a34925baf9

Observation fdc0672d-d3b9-42b4-bd0a-42fec5d3b1a5 · outbound

This paper cites arXiv preprint arXiv:2505.19092 (2025).

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation arXiv preprint arXiv:2505.19092 (2025)

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T10:06:30.236804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:6b25d26cd64c370a08e4af76cc4dc8813bc77c5ad1a961888988a2bec9e0f40a

Observation 986c789a-50c2-4e7e-b21b-4d7259389f65 · outbound

This paper cites Recommender systems with generative retrieval.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Recommender systems with generative retrieval

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.244072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:5f2631e1ae46781c3fb12618d7cfea5449447a1eed98fecab0a2604a8f121919

Observation f9af480d-8064-4dca-af9f-020a70c5992c · outbound

This paper cites ActionPiece: Contextually Tokenizing Action Sequences for Generative Recommendation.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation ActionPiece: Contextually Tokenizing Action Sequences for Generative Recommendation

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:30.248571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:6af05c09700f9718a152c972967836ee1e29f03d5981948f969f5d20329b8f45

Observation 03b7c38b-faee-4235-98e1-654675ea2aab · outbound

This paper cites Learnable item tokenization for generative recommendation.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Learnable item tokenization for generative recommendation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.292050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:4783d7c8167077fd535acbc331a8ca40fd7d53dc35c481c84c21c9e7e9369c7d

Observation 7c322a6f-11fe-4295-a86f-00111a1d82d7 · outbound

This paper cites How to index item ids for recommendation foundation models.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation How to index item ids for recommendation foundation models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.156875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:91c3ac711fe6e26caf27306540a8bc61e8048e2be03b5770a623b587a9a6dd84

Observation 53e4e53b-12f6-4c2e-b62e-3c9aefcf6478 · outbound

This paper cites Recommendation as language processing (rlp): A unified pretrain, personalized prompt & predict paradigm (p5).

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Recommendation as language processing (rlp): A unified pretrain, personalized prompt & predict paradigm (p5)

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.168794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:d901d4e8b0deed207be941b3e5016b3e2f9360a503b3de14979a4cebe30a4ebe

Observation 07bde36d-8eaa-48e6-b8f2-dba70125f70d · outbound

This paper cites Eager: Two-stream generative recommender with behavior-semantic collaboration.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Eager: Two-stream generative recommender with behavior-semantic collaboration

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.226066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:c22ccc28268c2b916fa9e74707ccea62a730ed6632bdb58be3fe3fe17524240d

Observation 65fc0211-49e9-4dc6-882c-e6a46f9cceb6 · outbound

This paper cites Actions Speak Louder than Words: Trillion-Parameter Sequential Transducers for Generative Recommendations.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Actions Speak Louder than Words: Trillion-Parameter Sequential Transducers for Generative Recommendations

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-13T19:39:33.082504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:fdc23624cbcae6b6c76811f3db85f8a7e2e2d2b65de11e59ce6238769fc90a72

Observation 86e4deea-82f2-4278-8c42-247ec6d51a15 · outbound

This paper cites Sprec: Self- play to debias llm-based recommendation.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Sprec: Self- play to debias llm-based recommendation

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.183633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:21c8b6611ba8527c29c94001e97a9465a1ad23b93707ff0be44b7e26b3fb1dcc

Observation e6205265-dad3-4e70-8255-a5af0d59e094 · outbound

This paper cites Process- supervised llm recommenders via flow-guided tuning.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Process- supervised llm recommenders via flow-guided tuning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.247454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:3ae8725a3c962d310df4e08a3cfce552ca6a57d382d59a9de55e6bcd3d7f2c9b

Observation 5a7de5c2-3b1c-49f2-8f5c-57ede7095d87 · outbound

This paper cites Nextquill: Causal preference modeling for enhancing llm personalization.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Nextquill: Causal preference modeling for enhancing llm personalization

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:30.218617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:3158233925b1274976b1d27ab7b64edf9094b747797841cd4a3e620f0c92d717

Observation cd0301b7-eee4-4b35-82b1-244d590e497d · outbound

This paper cites Don’t start over: A cost-effective framework for migrating personalized prompts between llms.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Don’t start over: A cost-effective framework for migrating personalized prompts between llms

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:30.224646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:7600ed5d51c66c8eb8f029b0acf090ecb455c1393b632307f3b36cf5a1c05509

Observation 1356707f-47fe-4ec8-8235-31a7201f256e · outbound

This paper cites Think-While-Generating: On-the-Fly Reasoning for Personalized Long-Form Generation.arXiv preprint arXiv:2512.06690.2025.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Think-While-Generating: On-the-Fly Reasoning for Personalized Long-Form Generation.arXiv preprint arXiv:2512.06690.2025

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:30.205546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:f1099478059d76415ab3841d35bfcdd7cbb04ec034ddbe5a86d2d10d62ab1271

Observation fd4e0302-9129-4df9-a5b5-13577fa35f70 · outbound

This paper cites Integrating large language models with reinforcement learning: A survey of llm-rl synergistic recommendation.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Integrating large language models with reinforcement learning: A survey of llm-rl synergistic recommendation

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.229947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:58eaf4b302bdf2f6a893f46590531eede3f5fb4219fe9e6a14af802a17fe2522

Observation 46f50316-c351-477b-b2cc-863897b32485 · outbound

This paper cites Generative retrieval with semantic tree-structured identifiers and contrastive learning.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Generative retrieval with semantic tree-structured identifiers and contrastive learning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.140971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:a5a25a1d8e57daba0e5f6057a67065fbdaef2ff57171b63d11c1fb8a7a39dbe1

Observation e20fa4b2-284a-4b30-b92b-f9ee65f9815e · outbound

This paper cites Idgenrec: Llm- recsys alignment with textual id learning.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Idgenrec: Llm- recsys alignment with textual id learning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.209302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:26d49804ed69e2d336aae69e22fa985b4a09354560d53f924856819fb5c4470e

Observation 732cbecb-e851-46ef-9708-c05e12ef2e1a · outbound

This paper cites Wide & deep learning for recommender systems.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Wide & deep learning for recommender systems

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.261750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:3ed3a702d4d53960eb718fb4b596ee0a1580bf6cfb2d5651a336f7a2516e6ff5

Observation 7ba00ccc-2175-4a4c-ab3b-af04067d801e · outbound

This paper cites Lightgcn: Simplifying and powering graph convolution network for recommenda- tion.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Lightgcn: Simplifying and powering graph convolution network for recommenda- tion

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.282832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:d1da7caf616c132e84d928ea53d0983c025bfd26f4447347028f6c4beb7aa556

Observation 126bc32d-07b8-4764-8e51-69ab8293032a · outbound

This paper cites Second order derivatives for network pruning: Optimal brain surgeon.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Second order derivatives for network pruning: Optimal brain surgeon

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.198930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:3a28464520be487bd290e63ee0181dc8157c70077c8ce2f225abc71e414cd26e

Observation 7d303553-74b8-48d9-898f-a269db93f897 · outbound

This paper cites Boosting Parameter Efficiency in LLM-Based Recommendation through Sophisticated Pruning.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Boosting Parameter Efficiency in LLM-Based Recommendation through Sophisticated Pruning

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:06:30.212147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:bcdc902d59c27389f7fe50744369cee9f7607922fb4fc482c1b8e143a3b4fb64

Observation 221e6220-3a08-49ed-96ac-bc232c7c8eac · outbound

This paper cites Smoothquant: Accurate and efficient post-training quantization for large language models.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Smoothquant: Accurate and efficient post-training quantization for large language models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.190671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:268dc6b33fa4673ec192881c55611684a2dbee3d0f278dbcf7d019cc4903a79e

Observation 5d4bd148-b2a7-452b-937d-40dc05deae7b · outbound

This paper cites Inductive generative recom- mendation via retrieval-based speculation.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Inductive generative recom- mendation via retrieval-based speculation

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.212673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:7efa8e3341aa518bcc7a823ae4e99f7fc357382d65acd41b69eda868112859b3

Observation 10456410-26dd-46ed-b2c9-15544ad62923 · outbound

This paper cites Nezha: A zero-sacrifice and hyperspeed decoding architecture for generative recommendations.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Nezha: A zero-sacrifice and hyperspeed decoding architecture for generative recommendations

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.233589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:42fb59b014f623cc86488f45ec4eda7df08e92226e05b188f7ff1627330610ce

Observation 037f22a4-d86c-4f6e-85f4-a79bebd3951e · outbound

This paper cites Earn: Efficient inference acceleration for llm-based generative recom- mendation by register tokens.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Earn: Efficient inference acceleration for llm-based generative recom- mendation by register tokens

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.216140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:c87ce087689cd9fdff8edaaef7f8ddf515754a7f159501824c1812a05637e2a6

Observation f3a0ddc1-07c3-4da0-b298-196fadc7d9d9 · outbound

This paper cites Justifying recommendations using distantly-labeled reviews and fine-grained aspects.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation Justifying recommendations using distantly-labeled reviews and fine-grained aspects

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.265337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:38b63e03dcea35373508ec354fafe48ddf7c0ef2ffdde532d7929b9d5a19a424

Observation 9120ed13-f557-4aeb-b7f5-bb7d5917b393 · outbound

This paper cites The llama 3 herd of models.

Position-Aware Drafting for Inference Acceleration in LLM-Based Generative List-Wise Recommendation The llama 3 herd of models

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T09:28:54.164808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T07:42:36.852261Z digest=sha256:90c8a3eaac227da2db9d0bf7502d6fc80ad4ac7f47ac5a6c25d3f99024fd776c

Pith citing papers

No inbound Pith citation observations are available.