Pith. sign in

Paper Citation Record · LEDGER

SpecLA: Efficient Speculative Decoding for Linear-Attention Models

As of 21 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2607.16673.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.16673 v1

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T20:19:46.943890Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 47197aa4-8472-44f3-b25f-02ede60ffc4c · outbound

This paper cites Hydra: Sequentially-Dependent Draft Heads for Medusa Decoding.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Hydra: Sequentially-Dependent Draft Heads for Medusa Decoding

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:44.468840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:44.468840Z digest=sha256:5b3369c034bfa2849053b74f120356dd8d841308078f61c7759ebaa548c55d51

Observation a9355325-7e93-4715-8e54-dce1dbe32e1c · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:44.570836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:44.570836Z digest=sha256:98ac1e28b3dfe785f5d57eb3ad71ae4e81309a5b5154d94c2963ca041031b70e

Observation ca9166a2-c1c1-4700-af34-ecf8e0a455d9 · outbound

This paper cites Lee, Deming Chen, and Tri Dao.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Lee, Deming Chen, and Tri Dao

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:44.795287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:44.795287Z digest=sha256:e01e1f4bf1fa952a627873a84c15d3ee1ee6b4843ba4f613791ee323568b95ba

Observation 74384af6-6074-4869-9dd8-4920296e18a6 · outbound

This paper cites Accelerating Large Language Model Decoding with Speculative Sampling.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Accelerating Large Language Model Decoding with Speculative Sampling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:44.864627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:44.864627Z digest=sha256:efcfa6a2b6a64318566b7eb6af48a377944389b2a9ada1169f1f05e57f143ccd

Observation b4b7416e-e9bf-4d7b-91f2-c5e059305b7b · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:44.969956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:44.969956Z digest=sha256:b60848ec4068f9933c5d66422d767538b79fa3045bd26470b206aad0553d1470

Observation ed31f3b6-7cc3-4979-a451-2b202ffd0f41 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:45.081917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:45.081917Z digest=sha256:38719cf96a0818a4ac067a43832338ff4602f95c449534bc22015a38e19d7885

Observation 73b0a4f6-d72f-4f4c-b47e-394cca338861 · outbound

This paper cites Prasanna.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Prasanna

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:45.164954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:45.164954Z digest=sha256:861c710e75059643fcc011ee2ddfcdee6aead288da75462148f0ae7856ee841b

Observation 94413e30-a693-435c-9038-c8d829e3db3e · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Sto- ica.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Gonzalez, Hao Zhang, and Ion Sto- ica

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:45.354547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:45.354547Z digest=sha256:cd67aac3755b2c82e50c1d9037d4736f97adc773b73888c862d4e5d31663ba59

Observation 31157899-1f90-4987-b73e-5935e8ade092 · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:45.448459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:45.448459Z digest=sha256:2bc95e43c00165d5136f264b51c59298e63d955f0f0a71bd392cce5b6e41dc4b

Observation 6cc75234-f5e2-47d6-b0d4-b9f1e7b013d6 · outbound

This paper cites EAGLE-2: Faster Inference of Language Models with Dynamic Draft Trees.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models EAGLE-2: Faster Inference of Language Models with Dynamic Draft Trees

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:45.490838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:45.490838Z digest=sha256:0847f2969e0ed41f3aacaa29f14a16641167e960cd860b7c0b2b52d6c1d7ae00

Observation 7220748c-6821-4677-9c8e-6f72fde2ba9e · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:45.565607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:45.565607Z digest=sha256:7d5dd887d780d177d2d6b19f6e28690ffea2e2336a42b99d7152801f024ebc70

Observation ed6bddaf-7008-42ae-b7e9-2200d18c2bb1 · outbound

This paper cites EAGLE-3: Scaling up Inference Acceleration of Large Language Models via Training-Time Test.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models EAGLE-3: Scaling up Inference Acceleration of Large Language Models via Training-Time Test

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:45.635216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:45.635216Z digest=sha256:eeefd57aac79f49efe6ae00a2beb15bc468257a4262864448ad56a271b0897c1

Observation 78ecd64d-95bb-4bd4-a59d-a12947607b8a · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:45.689578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:45.689578Z digest=sha256:a15c2a9b08acbf1dcd06176aa6441e90fd0cfdc7b1e7c4e70c85518aa5bd236c

Observation 35af792b-667d-4a70-b21c-3473c8ea45f4 · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:45.765931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:45.765931Z digest=sha256:85dee2e81540398d5b72628862eaa64d7cdf8eb108702581dc55ec9431c898d3

Observation 3c188c72-bf60-4ae6-97c1-fb10eb930f0c · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:45.833691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:45.833691Z digest=sha256:d9de17f6e7330f38a2c0c04f258eff4c08be0eee057153e80d183e39edd5b865

Observation b6bac746-3df4-4497-bb8b-5082113583f7 · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:45.881868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:45.881868Z digest=sha256:eba0b3aea999feeac1085e6af8a50477f519d6dffbeacace174c5d6cb04d8bd4

Observation 81cd479b-af67-4208-84b7-963f2ef20790 · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:46.024061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:46.024061Z digest=sha256:e0555403b55c618c5a484c02fa63c4bea44677d5e13e26b1f0da12acf1a9d62d

Observation ee090f6f-0795-4395-88c4-0dd02180dc43 · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:46.113630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:46.113630Z digest=sha256:b0cb2ba453424519ce6c3659a44e9131552cedb4285acac6d0407ba985fe514b

Observation 386fd394-7545-482e-8937-413b2559c1f8 · outbound

This paper cites Qwen3 Technical Report.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Qwen3 Technical Report

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:46.203786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:46.203786Z digest=sha256:412422434d707fdd57e725e93e47863b48dca0a58e6716e65033ad256ced1ec0

Observation 436c64e6-b492-4052-b755-58e359649f93 · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:46.270209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:46.270209Z digest=sha256:8c99af053905daad306ec45c6561caaefdffc9786a41a6f3a6168a1bedf6cbd8

Observation a80fe95e-a18e-458a-a047-67b37a11d1b1 · outbound

This paper cites Gomez, Lukasz Kaiser, and Illia Polosukhin.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Gomez, Lukasz Kaiser, and Illia Polosukhin

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:46.360892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:46.360892Z digest=sha256:7ea6b80bc82f6807589aac4f152b1003c801d62e76e294565fe298ce5fc0f703

Observation 9583f771-ba9c-4c03-b663-ecf2a5e74d92 · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:46.436936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:46.436936Z digest=sha256:11c8e96485d42666a5cfe166b89f51c14de8c96530457e013ae96163bb75be64

Observation fe7d6a8e-275f-4058-bd69-92e457af3878 · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:46.501977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:46.501977Z digest=sha256:1f3e914b38a88a1fc561a6108b3d0a0432949c476c59cfb8d5741338dd763697

Observation 8035a39e-49b4-427f-a836-aeac380d5954 · outbound

This paper cites Gated Delta Networks: Improving Mamba2 with Delta Rule.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Gated Delta Networks: Improving Mamba2 with Delta Rule

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:46.593633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:46.593633Z digest=sha256:38df9072e454c8a42f9da9111fc960cdfb6315c86b8fa96c7347986d5f60b17a

Observation 3f67ae79-a6aa-4c6f-ab50-eab0532cf842 · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:46.690162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:46.690162Z digest=sha256:2976cfece330811c96fe5fd45c0ea6f5e9ebd04addd3d3aeea5bba856be8b412

Observation b838c0bd-aa02-48ea-b6f9-dbe30dbcbd4e · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:46.760226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:46.760226Z digest=sha256:20e2f022f1b5dfa1db209d072e555f7e679b0963d6aa8c722b6186bfd6e5a68a

Observation 6b3b4eab-891c-4cfa-93c5-57b014966e9e · outbound

This paper cites an unresolved cited work.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:46.943890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:46.943890Z digest=sha256:1126158e98071d8e35cd1843bf259ee3d10a7893692def4607d8b97132f4a22a

Observation 16d3ec25-d680-41d7-b7e8-085ab29b01d2 · outbound

This paper cites InAdvances in Neural Information Processing Systems, Vol.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models InAdvances in Neural Information Processing Systems, Vol

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:45.953747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:45.953747Z digest=sha256:a1a62797b95c9332e58affa0537edba9454a696d8b9faf0bb78aa0e7d315de82

Observation 4bf6e88e-6e34-4ae0-b3b4-76fb566e3f1f · outbound

This paper cites InAdvances in Neural Information Processing Sys- tems, Vol.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models InAdvances in Neural Information Processing Sys- tems, Vol

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:46.860632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:46.860632Z digest=sha256:24bebe3ab9d5b02d17c1bd9576c7d8c117c8cace9c89c46913503e0d9efe4888

Observation 56fa3ec7-bc09-4ce8-941b-63b7b7eda8bf · outbound

This paper cites arXiv preprint arXiv:2503.14376.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models arXiv preprint arXiv:2503.14376

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:44.701711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:44.701711Z digest=sha256:12aa54c659d752b6f5738152439437e2396aefa32d4a2f2194273f6df937e60a

Observation f0864df3-7a2b-476b-9a5a-de8c17440a56 · outbound

This paper cites arXiv preprint arXiv:2603.05931.

SpecLA: Efficient Speculative Decoding for Linear-Attention Models arXiv preprint arXiv:2603.05931

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-01T20:19:45.246592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:19:45.246592Z digest=sha256:eabb02132326dc912478eaf19acc1d2facfe9ff775d914affb6191844286fb68

Pith citing papers

No inbound Pith citation observations are available.