Pith. sign in

Paper Citation Record · LEDGER

How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2404.02690.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.02690 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T16:14:05.607282Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T14:43:22.563022Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6b5812cf-0fcc-4015-8a02-991f2e04b931 · inbound

RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval cites this paper.

RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 61

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T08:12:02.011756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-18T08:12:01.798459Z digest=sha256:fd72fef3881196b04dd2a142805265fa1c0c2058067cdfb50f1a83350a50069b

Observation 271fafd0-d81b-455e-a805-d2108fa7a407 · inbound

SCBench: A KV Cache-Centric Analysis of Long-Context Methods cites this paper.

SCBench: A KV Cache-Centric Analysis of Long-Context Methods How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T16:14:05.607282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T16:14:05.607282Z digest=sha256:8ff709af1d14c0ef8b31841ad4d7f5243b12e77e654579b8be470e3fa25c5f54

Observation 7efb6ab4-863e-4cb4-8da5-0ff89a75f6ff · inbound

HashAttention: Semantic Sparsity for Faster Inference cites this paper.

HashAttention: Semantic Sparsity for Faster Inference How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T12:18:19.392692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T12:18:19.392692Z digest=sha256:fc8efbb4bac503165a1bf5e06c941c2063972c3a52a4ae0406cca1b254af024d

Observation 5f031627-7f06-4564-8166-0c1a0750901d · inbound

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning cites this paper.

LeMo: Enabling LEss Token Involvement for MOre Context Fine-tuning How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T20:26:11.017792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:26:11.017792Z digest=sha256:bea79dfc27303d98a5cda41e320a73e9c71a0d274ec1de87925c73fd159a129a

Observation 6120744c-586e-439d-a8fd-b7678f16d20e · inbound

Video Latent Flow Matching: Optimal Polynomial Projections for Video Interpolation and Extrapolation cites this paper.

Video Latent Flow Matching: Optimal Polynomial Projections for Video Interpolation and Extrapolation How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T18:51:12.401249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:51:12.401249Z digest=sha256:8ebfe4632fb0e5cb76d97a0311f49e1317a91ba6a9f34b1d76728e7464861d82

Observation 306b84cf-2d37-462e-b28a-8c935a3f14ce · inbound

RetroInfer: A Vector Storage Engine for Scalable Long-Context LLM Inference cites this paper.

RetroInfer: A Vector Storage Engine for Scalable Long-Context LLM Inference How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:35:25.134551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-22T15:59:04.724780Z digest=sha256:8b84a85d3fa52a56412f356283b0d47b65f91fa010dba44127a86470e8b843de

Observation 9b8de900-60eb-4bab-8ac4-71d127049db3 · inbound

SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling cites this paper.

SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:03.324221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:03.324221Z digest=sha256:c25633f304beed5b013cdee4df74b793573012a9ee8d36dc68772476c23d3256

Observation a6c6c0d2-b311-4656-a2ea-ccd32cfdedfb · inbound

EARN: Efficient Inference Acceleration for LLM-based Generative Recommendation by Register Tokens cites this paper.

EARN: Efficient Inference Acceleration for LLM-based Generative Recommendation by Register Tokens How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:18:46.851223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:18:46.851223Z digest=sha256:956b3f60517b0a1dd1b72e69273625e95a95126c61b6fd7ab04294143aa03c2d

Observation eff45d7b-c6ee-48fb-9095-99acb06129ed · inbound

DeltaLLM: A Training-Free Framework Exploiting Temporal Sparsity for Efficient Edge LLM Inference cites this paper.

DeltaLLM: A Training-Free Framework Exploiting Temporal Sparsity for Efficient Edge LLM Inference How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T14:17:44.221622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:17:44.221622Z digest=sha256:2b7879909428a04c74f63963fc35cb32bc78d778ebc74bf81a60f6f1647fdfdc

Observation 0bfba7ea-7231-457d-8cc2-e55107447ccd · inbound

Unifying Learning Dynamics and Generalization in Transformers Scaling Law cites this paper.

Unifying Learning Dynamics and Generalization in Transformers Scaling Law How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T14:02:56.657443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:02:56.657443Z digest=sha256:3b4d18ee9d231136f65580a114b0bc17e0022ea9af5b6c90e5776153f9280ed6

Observation 682b3543-8763-4ac2-ac00-c0bb4ecce3cd · inbound

AdaCluster: Adaptive Query-Key Clustering for Sparse Attention in Video Generation cites this paper.

AdaCluster: Adaptive Query-Key Clustering for Sparse Attention in Video Generation How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:43:49.771996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T05:09:46.155328Z digest=sha256:f1a99e179633d79dedd295f00f29d5c2801365a04b330e0f2754e28bfe1f8351

Observation b8f0bd32-a0d4-457d-8005-c787e396ee0c · inbound

Unifying Sparse Attention with Hierarchical Memory for Scalable Long-Context LLM Serving cites this paper.

Unifying Sparse Attention with Hierarchical Memory for Scalable Long-Context LLM Serving How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:01:25.970065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-07T13:15:21.201950Z digest=sha256:26998972773b9155921398141c0ca9222b208c724c1c7ff8597b1d251dfaa090

Observation 7bd1e6bb-a2e3-4e1f-aade-e22b7ce1ed6c · inbound

Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction cites this paper.

Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:41:26.694598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-12T05:02:25.513351Z digest=sha256:f9add7334f23171532e473536aae248e56787a2dd6d9e4e34c5288667f490918

Observation 270bd426-51b4-4746-9629-c86ccca39f24 · inbound

HierEdit: Region-Aware Hierarchical Diffusion for Efficient High-Resolution Editing cites this paper.

HierEdit: Region-Aware Hierarchical Diffusion for Efficient High-Resolution Editing How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:43:22.565139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T14:40:20.517450Z digest=sha256:4104ad34b9f725f2632da595d623e9d598187e4ce859107de274dda535abd92f

Observation 788cdefd-0f10-471b-b781-cc76670f2e2d · inbound

SPADE: An Input-Adaptive Sparse Attention Engine for Fast Video Diffusion Models Inference cites this paper.

SPADE: An Input-Adaptive Sparse Attention Engine for Fast Video Diffusion Models Inference How Sparse Attention Approximates Exact Attention? Your Attention is Naturally $n^C$-Sparse

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T20:50:03.960063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:50:03.960063Z digest=sha256:cb4ee78a561c4a68f4d55a730319f043c6a3f99a26e752cd4805d8fd6ace2a25