Pith. sign in

Paper Citation Record · LEDGER

SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2406.15486.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.15486 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T19:52:04.194443Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4c542d0c-68fc-441d-bfee-5052355a5a64 · inbound

Political-LLM: Large Language Models in Political Science cites this paper.

Political-LLM: Large Language Models in Political Science SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-11T19:52:04.194443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:52:04.194443Z digest=sha256:0f9d0611d653055474b4be7bc58830a3595acfb4d5521bc1de1d5c021c29717d

Observation 9918342e-1efe-4d15-9373-8aab1085b533 · inbound

SCBench: A KV Cache-Centric Analysis of Long-Context Methods cites this paper.

SCBench: A KV Cache-Centric Analysis of Long-Context Methods SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-11T16:14:06.318876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T16:14:06.318876Z digest=sha256:044903358a3286cfdbc3aeb3deede1562a36e617dba911b6937b59b6367ff06f

Observation fe622c78-ea51-4154-a854-9c00766ff892 · inbound

SepLLM: Accelerate Large Language Models by Compressing One Segment into One Separator cites this paper.

SepLLM: Accelerate Large Language Models by Compressing One Segment into One Separator SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-11T14:25:06.112967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:25:06.112967Z digest=sha256:04918d5c7e921b368b74264056c6c82216d8b8a660bfbf1806e5d7bfc5bc8783

Observation e00000cd-e7e5-4938-b067-d560714c60f6 · inbound

Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation cites this paper.

Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-09T11:11:17.839813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:11:17.839813Z digest=sha256:325e0ca2e3fd1f80da7de61ceb9d5d37c562abb5d14be0a265a82f03fc3b95f7

Observation 880c0f54-d757-48f2-b888-07afa42268bf · inbound

SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling cites this paper.

SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:04.130346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:04.130346Z digest=sha256:d33bd3945ecfeaf9ff98cafe82bda7a805c64f0a3e04d7ebf102f445c3390efa

Observation 1e48076b-dbe4-4a8a-82fc-b0250b397dc8 · inbound

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU cites this paper.

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:11.244159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:00:11.244159Z digest=sha256:8bdf283eeff85a9f59c08eff50473e96865e1d84795112aab9d1c95b54b81391

Observation 11eee389-6bd1-4de9-84ad-83e1ceebfab3 · inbound

Unveiling Effective In-Context Configurations for Image Captioning: An External & Internal Analysis cites this paper.

Unveiling Effective In-Context Configurations for Image Captioning: An External & Internal Analysis SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:13.532463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:13.532463Z digest=sha256:5d3880b9c26fd67b68a7a163d428b6f690dd60bfb30b17661f45a29aa6b1de34

Observation f07f11f8-9fad-4f27-8450-67999036c943 · inbound

vAttention: Verified Sparse Attention cites this paper.

vAttention: Verified Sparse Attention SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T11:21:10.344773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:21:10.344773Z digest=sha256:c1ca8db25d4c4ce6e81b035e6e22a11e4106e4329f0911cc8099718e786a8e42

Observation e129d356-e531-4ea7-b65f-203fa9d92b89 · inbound

CSAttention: Centroid-Scoring Attention for Accelerating LLM Inference cites this paper.

CSAttention: Centroid-Scoring Attention for Accelerating LLM Inference SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:53:02.152558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-14T21:52:25.148104Z digest=sha256:d629793672f26e21955d27f7a3c935290fd896dab7a964cdba06b9221fbde124

Observation fd98a068-3906-4903-9d7c-a58fd7ffb68b · inbound

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving cites this paper.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:38:23.097181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:ce09d0453fa84f972cf13d47be19b7a502327ed3ae7f146adb609f779777eb20

Observation 736369e9-1bae-4162-a0af-46f3c09eb51d · inbound

Sparse Attention as a Range Searching Problem: Towards an Inference-Efficient Index for KV Cache cites this paper.

Sparse Attention as a Range Searching Problem: Towards an Inference-Efficient Index for KV Cache SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:15:50.583185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-11T01:15:35.871863Z digest=sha256:2243c03b161cca4b97f8ae2b8da382da57e460b1c5bc8579108590e4c3843196

Observation 502066ef-d2a9-4853-8c26-1d7d1db5d78a · inbound

SIFT: Selective-Index For Fast Compute of RAG Prefill by Exploiting Attention Invariance cites this paper.

SIFT: Selective-Index For Fast Compute of RAG Prefill by Exploiting Attention Invariance SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:37:31.226802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T16:24:31.109508Z digest=sha256:aacc470d49b7f8c0a1adda4b56d50265b704675fbcc7260f2bb0f75bce9d2d9d