Pith. sign in

Paper Citation Record · LEDGER

SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2406.15486.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.15486 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:39:04.130346Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 880c0f54-d757-48f2-b888-07afa42268bf · inbound

SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling cites this paper.

SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:04.130346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:04.130346Z digest=sha256:9327af75c214a7ae5a3a0e361076df15ef8a34803f7a4112682a1a04cf8051b8

Observation 1e48076b-dbe4-4a8a-82fc-b0250b397dc8 · inbound

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU cites this paper.

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:11.244159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:00:11.244159Z digest=sha256:ceb223b420cfa4ec3cd9f0d2fe7f5bcf732777637b7192922e70f14c13979d5e

Observation 11eee389-6bd1-4de9-84ad-83e1ceebfab3 · inbound

Unveiling Effective In-Context Configurations for Image Captioning: An External & Internal Analysis cites this paper.

Unveiling Effective In-Context Configurations for Image Captioning: An External & Internal Analysis SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:13.532463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:13.532463Z digest=sha256:45d378548f78b5b5acf0b328d1df2f5d970ad84ee3847728643864ee1ec8ff78

Observation f07f11f8-9fad-4f27-8450-67999036c943 · inbound

vAttention: Verified Sparse Attention cites this paper.

vAttention: Verified Sparse Attention SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T11:21:10.344773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:21:10.344773Z digest=sha256:d625973395a02913115c927dca3a4045d9607fd69ac675850944a500796c5987

Observation e129d356-e531-4ea7-b65f-203fa9d92b89 · inbound

CSAttention: Centroid-Scoring Attention for Accelerating LLM Inference cites this paper.

CSAttention: Centroid-Scoring Attention for Accelerating LLM Inference SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:53:02.152558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T21:52:25.148104Z digest=sha256:88ae5215813a00a040430d8ff19727e5866b532c993855d6ba3c435daaaab8af

Observation fd98a068-3906-4903-9d7c-a58fd7ffb68b · inbound

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving cites this paper.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:38:23.097181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:c38eaf6c6c6a303b015de7d0f56d40ca5d977c881b3a451565942c3e00279bcb

Observation 736369e9-1bae-4162-a0af-46f3c09eb51d · inbound

Sparse Attention as a Range Searching Problem: Towards an Inference-Efficient Index for KV Cache cites this paper.

Sparse Attention as a Range Searching Problem: Towards an Inference-Efficient Index for KV Cache SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:15:50.583185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T01:15:35.871863Z digest=sha256:59be04a6ddfad435d29295560aeae9d16ea03ddc06285fa9069eefbcef9ce49e

Observation 502066ef-d2a9-4853-8c26-1d7d1db5d78a · inbound

SIFT: Selective-Index For Fast Compute of RAG Prefill by Exploiting Attention Invariance cites this paper.

SIFT: Selective-Index For Fast Compute of RAG Prefill by Exploiting Attention Invariance SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:37:31.226802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T16:24:31.109508Z digest=sha256:620cfb62ba1fb32a8acaa817476b85276f5c09c1d1b205fbd48302bc5f219fd4