Pith. sign in

Paper Citation Record · LEDGER

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference

As of 8 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2511.21702.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2511.21702 v2

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T22:05:05.627076Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f4a3b168-a307-498c-ab29-7cdf42eea5df · outbound

This paper cites Attention is all you need,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Attention is all you need,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.591222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.591222Z digest=sha256:61dabb6c5e6ae6071899f407105e82b3035bb5599f935d258d5833d4ee3b9a68

Observation e40bd281-0255-4eca-9c3c-add394adf813 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.594492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.594492Z digest=sha256:dec70b7a6b3c76f9237872681279910a4dd3bbb79f4b7db9a3ff39bd84474e3e

Observation 9da969fa-b914-4b1e-b3d8-fcd03967ae2e · outbound

This paper cites Mistral 7B.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Mistral 7B

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.597494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.597494Z digest=sha256:f39db20c8b236bb1a6460d7e0a1b4dba0382a6d8bf43aea5fb046a9dda783457

Observation 53c1ac62-143c-47c2-88b2-471c4cb9174c · outbound

This paper cites Efficient softmax ap- proximation for gpus,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Efficient softmax ap- proximation for gpus,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.600504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.600504Z digest=sha256:e2bcb84bc864ee7bee7ab8702292c71fd520de2d4664f1cb58f200f40b4468ed

Observation 72ab3e34-582a-4c24-b05c-45f3ee453e17 · outbound

This paper cites Efficient estimation of word representations in vector space,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Efficient estimation of word representations in vector space,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.603392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.603392Z digest=sha256:53edcd9dbe150ded9bd6b5d8613a11188b1723aadf7d799bf8cf3370ea677024

Observation 357b694e-a262-45b9-8bd0-84ddad0d6793 · outbound

This paper cites Sparse convolutional neural networks,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Sparse convolutional neural networks,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.606140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.606140Z digest=sha256:0a4a49589d71ac31946fa122e07a6bdb691ca989503302bc8e9e0abc4f02b542

Observation f341e3fd-4c5f-44c2-8424-18c99136736d · outbound

This paper cites Reducing transformer depth on demand with structured dropout,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Reducing transformer depth on demand with structured dropout,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.609010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.609010Z digest=sha256:6f8560f80f86ea1b27affaa8aa07c8748905514fdc0a7da51b6799c392e58a5d

Observation 994039a7-02ed-4334-82c2-5372deded919 · outbound

This paper cites Fast inference from trans- formers via speculative decoding,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Fast inference from trans- formers via speculative decoding,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.611417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.611417Z digest=sha256:c11fd67c4b7632e936035826b02ebb57c61e5f048b8d652abb2aea0645d8636f

Observation 18f5aabe-e6a3-4ec4-98fe-390be31f5dde · outbound

This paper cites Designing Large Foundation Models for Efficient Training and Inference: A Survey.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Designing Large Foundation Models for Efficient Training and Inference: A Survey

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.613847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.613847Z digest=sha256:b40e3798d4e6bab1f83a0ff714f1ae3c1b3b2f4490ce51eea14fb5ac876a97b6

Observation 81613c19-4385-4880-919b-d612802bda55 · outbound

This paper cites MKA: Memory-keyed attention for efficient long-context reasoning,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference MKA: Memory-keyed attention for efficient long-context reasoning,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.616895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.616895Z digest=sha256:5196bf8fb40840f30b3118df6cb300ef2403bc4f4b7e88e8dda14e596b75754b

Observation 9fb5c40e-6a18-41c8-b9e5-9d213d7c4ba3 · outbound

This paper cites Tinyserve: Query-aware cache selection for efficient llm serving,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Tinyserve: Query-aware cache selection for efficient llm serving,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.619389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.619389Z digest=sha256:f8d664c955b58c998ad13f650eadcbb2795a357da331ddc9f7733ce47da16c9c

Observation b9d4f2f3-e95f-4ce3-a933-9f1762b992c7 · outbound

This paper cites PiKV: KV Cache Management System for Mixture of Experts.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference PiKV: KV Cache Management System for Mixture of Experts

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.621861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.621861Z digest=sha256:3ca1c87f2cb60c38901f718b6a53cdf6075dd09f8ed602fcbb69e0d751c73c12

Observation 4f9a152f-3f27-4956-b755-346cfcb69f41 · outbound

This paper cites Llmeasyquant: Scalable quantization for parallel and distributed llm inference,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Llmeasyquant: Scalable quantization for parallel and distributed llm inference,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.624658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.624658Z digest=sha256:d07d11a86be203160ea935ad7d5a415d0a14b94fce052949a26c55de8e6492fa

Observation 9036fda3-d7e0-4d49-8dd1-24fd1d3ae901 · outbound

This paper cites Flashattention- 2: Faster attention with better parallelism and work partitioning,.

CSV-Decode: Certifiable Sub-Vocabulary Decoding for Efficient Large Language Model Inference Flashattention- 2: Faster attention with better parallelism and work partitioning,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T22:05:05.627076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:05:05.627076Z digest=sha256:6688be6f08215d9de448b9791ccd651d7e5aafb95fe2bbeb8ab7820c9f19dbcf

Pith citing papers

No inbound Pith citation observations are available.