Pith. sign in

Paper Citation Record · LEDGER

Loki: Low-rank Keys for Efficient Sparse Attention

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2406.02542.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.02542 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:29:13.910553Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8ca35de0-fc3f-449f-8684-e8f0c36c06d9 · inbound

RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval cites this paper.

RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval Loki: Low-rank Keys for Efficient Sparse Attention

Reference 108

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:12:01.907754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T08:12:01.798459Z digest=sha256:81e049b1b4de55494cd314667ec13a4f45cf28c38483adff52e2fb580da5678d

Observation 4930a632-4f1d-42cd-b2d2-53356ebee0ba · inbound

Exploiting Sparsity for Long Context Inference: Million Token Contexts on Commodity GPUs cites this paper.

Exploiting Sparsity for Long Context Inference: Million Token Contexts on Commodity GPUs Loki: Low-rank Keys for Efficient Sparse Attention

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T14:29:13.910553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:29:13.910553Z digest=sha256:467d7ad14d460375d6a90a99f7bcc83047733bdb6beeb99cdf18f2608cfdd7da

Observation 5adf0400-21db-4fc6-8c8a-0c87a08d3785 · inbound

Hardware-Efficient Attention for Fast Decoding cites this paper.

Hardware-Efficient Attention for Fast Decoding Loki: Low-rank Keys for Efficient Sparse Attention

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T13:32:32.523852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:32:32.523852Z digest=sha256:74022124cc546d11a644759ca41d0fa4b2771f5d77a54d667c1efea86410e045

Observation db9cac96-1645-4357-b4db-31300544deb0 · inbound

HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference cites this paper.

HATA: Trainable and Hardware-Efficient Hash-Aware Top-k Attention for Scalable Large Model Inference Loki: Low-rank Keys for Efficient Sparse Attention

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:00.797982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:26:00.797982Z digest=sha256:4440c4f393017a7fca55f80aa3efb36bb3bd6b14c9a9eac6d04cf9a125d3ca09

Observation fa23b0ac-9a7d-4eeb-9982-bf2e55b301c7 · inbound

GraphKV: Breaking the Static Selection Paradigm with Graph-Based KV Cache Eviction cites this paper.

GraphKV: Breaking the Static Selection Paradigm with Graph-Based KV Cache Eviction Loki: Low-rank Keys for Efficient Sparse Attention

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T13:44:18.504592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:44:18.504592Z digest=sha256:3c22b1b214415941e5d8000c818802f6edbd5029defca33556cd44999d3996f7

Observation 49f7425a-66a5-403d-bdf2-06b232b5d774 · inbound

AQUA: Attention via QUery mAgnitudes for Memory and Compute Efficient Inference in LLMs cites this paper.

AQUA: Attention via QUery mAgnitudes for Memory and Compute Efficient Inference in LLMs Loki: Low-rank Keys for Efficient Sparse Attention

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T17:05:55.021776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:05:55.021776Z digest=sha256:f79d59431f53d2177d9a4229ccea10ee07ff9c824fe6f342e7ec1f87ce0eec41

Observation ca6abd9e-09b2-4c1c-9e38-d7e797e76c40 · inbound

vAttention: Verified Sparse Attention cites this paper.

vAttention: Verified Sparse Attention Loki: Low-rank Keys for Efficient Sparse Attention

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T11:21:08.754040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:21:08.754040Z digest=sha256:1d5801a803b3538e9c6cf11e8fa6fa4462499ba621af70f8655d48084e5b6b6c

Observation 3e6c4aed-007c-465a-9860-6dc2b669f48b · inbound

Why Attend to Everything? Focus is the Key cites this paper.

Why Attend to Everything? Focus is the Key Loki: Low-rank Keys for Efficient Sparse Attention

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-15T11:59:59.625594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-15T11:55:36.701888Z digest=sha256:8076c2de92f59043d3fd52b4e85de282547f54f163ba87b459a3f06bc7d8debc

Observation 84050e28-e812-4a6d-9aeb-ab6d42359d72 · inbound

HieraSparse: Hierarchical Semi-Structured Sparse KV Attention cites this paper.

HieraSparse: Hierarchical Semi-Structured Sparse KV Attention Loki: Low-rank Keys for Efficient Sparse Attention

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:16:54.233089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T07:15:19.184970Z digest=sha256:cb9ed0592a117befbe440d7dd87189629485e13f5602f660a490f965bb7ec481

Observation c9989967-2c6e-4f3c-9647-7427c7307e77 · inbound

Sparse Attention as a Range Searching Problem: Towards an Inference-Efficient Index for KV Cache cites this paper.

Sparse Attention as a Range Searching Problem: Towards an Inference-Efficient Index for KV Cache Loki: Low-rank Keys for Efficient Sparse Attention

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:15:50.565136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:15:35.871863Z digest=sha256:18e631abf6448c976600c4e68770aca0338f7cbc7c26ad9ce39c52b629aa1f0d

Observation d7c83128-372f-40bf-9acc-9130da661a36 · inbound

COBS: Cumulant Order Block Sparse Attention cites this paper.

COBS: Cumulant Order Block Sparse Attention Loki: Low-rank Keys for Efficient Sparse Attention

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-13T00:42:31.008284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T00:42:31.008284Z digest=sha256:8954cf00eedcee2f11022082ca3fc5f20a629e4ed7bb7dc8063032b4c481d742

Observation 5393c3b5-e078-4c43-8730-3640dc0c7287 · inbound

LOCKS: Page-Local Compact Key Summaries for Efficient Long-Context Decoding cites this paper.

LOCKS: Page-Local Compact Key Summaries for Efficient Long-Context Decoding Loki: Low-rank Keys for Efficient Sparse Attention

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-31T11:49:11.732389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T11:49:11.732389Z digest=sha256:1b10394018caba5c7b6d1eca6655cc05409ecd116278609f1b6bc63d8b416d39