Pith. sign in

Paper Citation Record · LEDGER

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing

As of 15 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 3 inbound Pith citation observations for arXiv:2505.19578.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19578 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:16:40.355183Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-23T03:59:58.634512Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T04:02:30.324254Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1dc9b640-3ba3-49b7-98d6-99fe94da9097 · outbound

This paper cites online" 'onlinestring :=.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.469493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.469493Z digest=sha256:6aa5bb9dc329c301b006049e05adac0fc53700c7f65fb20a480013b350a4304e

Observation 827d8313-2989-4717-a290-4af0085fbbd4 · outbound

This paper cites write newline.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.543759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.543759Z digest=sha256:827eba036dc3514d55ace4eb18bc31070bed80a006684a75fecb07234c1c50c1

Observation 6915cbc0-5938-4780-b9e4-470887c3aa27 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:42.431606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:16:38.615637Z digest=sha256:abe255367fcb7086a71445c966eb19969a0e5ea12cc6b6fd3452dc095f2082a4

Observation 213b999d-51af-45e2-b971-ca71fd5341a7 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:42.237866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:16:38.682780Z digest=sha256:59cc7051fed55175239911f7a29f1508de5616647100abbeb47f918b0b75d876

Observation e4159931-1ccb-400a-bba9-23c45a77975f · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:42.027748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:16:38.747418Z digest=sha256:06d07dda17155509d5433635f9b24f464080f2cc1eed0a79869c6a5fda5c97f4

Observation bf3a50cf-0d8b-4ffe-b341-7bea06ba24ca · outbound

This paper cites Challenges in Deploying Long-Context Transformers: A Theoretical Peak Performance Analysis.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Challenges in Deploying Long-Context Transformers: A Theoretical Peak Performance Analysis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.844275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.844275Z digest=sha256:fafdb575cd3bc3666a69fe6466e91bdb39ad8c4b58fb02dfe46d2dfc21025144

Observation 3dc423ee-3c08-4c77-8182-bfee253b9632 · outbound

This paper cites Chain-of-Thought Hub: A Continuous Effort to Measure Large Language Models' Reasoning Performance.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Chain-of-Thought Hub: A Continuous Effort to Measure Large Language Models' Reasoning Performance

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.906618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.906618Z digest=sha256:6bb7181096b0437f554634924228c95a50076d3c939826cc393233df4cea3779

Observation 1338f882-ee10-40dc-8718-847d1628ef14 · outbound

This paper cites SeerAttention: Learning Intrinsic Sparse Attention in Your LLMs.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing SeerAttention: Learning Intrinsic Sparse Attention in Your LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.982877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.982877Z digest=sha256:eb39ce0da87181b2c225d08bc5c702309c36df8e594727a22587ecc8c6c72d36

Observation 6c284909-fc44-408e-8529-3d4c55203ece · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:41.825903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.069843Z digest=sha256:1971e11c6169a79c3c86eaa22e22598bea49efb5174eb35a91f3dc29893c758a

Observation 9e4143fe-b934-46ba-a450-5ca26e49ac93 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:41.593628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.137625Z digest=sha256:bc234c1556dd691f90146693743d1d814aa5448a7e86a98ccb15efcc5ed9cffc

Observation ba716828-7304-4818-a5b0-98e495fb9fbd · outbound

This paper cites MoBA: Mixture of Block Attention for Long-Context LLMs.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing MoBA: Mixture of Block Attention for Long-Context LLMs

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.237793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.237793Z digest=sha256:eac4226ae00dedeb7350afe95c2ebb1bc5db4cac4cd814a6856cea4a7f4b70b3

Observation a840e619-b27d-4529-a078-8c29998a3e3d · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:41.433549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.305185Z digest=sha256:86d14493a84a138c365b0aeb77c016c082b047a272de8afa5e801b280a984918

Observation 7fe57c19-3de6-4f68-8951-80766f1a1c2a · outbound

This paper cites Rae, Anna Potapenko, Siddhant M.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Rae, Anna Potapenko, Siddhant M

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:41.256030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.374466Z digest=sha256:ec71d0ae8112f2d5f7e2187936beb8e32ea0a74b13a6a5c65604bfff0022a62c

Observation ba8af17f-dfd1-4ac8-b226-bc342feab119 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.492504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.492504Z digest=sha256:923b17ee85fadec4a78e95ff08d5f53a8a2a420a1e8cada95c9fc18081c20498

Observation 156a87a1-6eee-47df-8e8b-6c232526fe15 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.578526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.578526Z digest=sha256:85c61fb5ac09e2f83b4c921155f6589e1339beb7e0dcc6d4c4e940f1e7a08991

Observation d1cb9a39-2d09-4316-ba70-0d18eee7356d · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.677520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.677520Z digest=sha256:cecb512e4269182d132d08bbdcd4d85b862f4b8d9c34252a4a84825d662d23c6

Observation f17f4dc1-b674-4621-81e6-37dd411cba22 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:41.053429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.816195Z digest=sha256:96f4762927de3c2f822dcb8cd0d34a7727f3d88401339009c43cb97333ec991d

Observation 86afec40-3f91-4afe-9d08-80d28665fd63 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:40.874801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.924333Z digest=sha256:df89ec30e133543d9e422bc5bcc32058fc584b8ed5999063d32ff89b73957369

Observation 0c668210-0348-4e32-97f0-15d24d57d5c1 · outbound

This paper cites Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.017957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.017957Z digest=sha256:aa51f653a25e5a43468f47cd4a4258b35bc5de027b74bd3842b4fa55b6695535

Observation 153761b0-0eff-4f28-99c2-a3794f7c0422 · outbound

This paper cites Multi-agent Architecture Search via Agentic Supernet.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Multi-agent Architecture Search via Agentic Supernet

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.125801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.125801Z digest=sha256:1183de39b3eefe15141cd154c74ed4e558c670a1d3c92b20f4ea9f59601c2af8

Observation 426a03a4-1ef5-42d8-a68b-01905b6e7ef9 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:40.676404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T14:16:40.247475Z digest=sha256:be7366ab1ce226f13183da47c3ad6bf1558ec1df162719c948f32822bcfedf0c

Observation cfd5149f-e717-47e3-b0a7-3d786cb52dd1 · outbound

This paper cites Migrating Code At Scale With LLMs At Google.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Migrating Code At Scale With LLMs At Google

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.355183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.355183Z digest=sha256:f101360e70862ecbdf32b636560ee308f7f2e6ab52ae1e9a8cce9d19c1fcf4e5

Pith citing papers

Observation 39dad653-1567-44cb-8789-578909a41159 · inbound

FastKV: Decoupling of Context Reduction and KV Cache Compression for Prefill-Decoding Acceleration cites this paper.

FastKV: Decoupling of Context Reduction and KV Cache Compression for Prefill-Decoding Acceleration Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:02:30.327548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-23T03:59:58.634512Z digest=sha256:4dc005ad787495c57142107dd6219ba9a016cf0f8d3d71007c67b5a0a9d17e15

Observation 80c0fe47-b306-47a8-bb3c-4cb78229bbed · inbound

EchoKV: Efficient KV Cache Compression via Similarity-Based Reconstruction cites this paper.

EchoKV: Efficient KV Cache Compression via Similarity-Based Reconstruction Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:09:36.929964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-15T01:09:07.983785Z digest=sha256:b051bc5d45de60596c51be1bb1079b60fc3d8161bfc0eb230215b66d313de214

Observation 7637f781-4ef2-450b-b8d1-c632a0af0ef3 · inbound

Flux Attention: Context-Aware Hybrid Attention for Efficient LLMs Inference cites this paper.

Flux Attention: Context-Aware Hybrid Attention for Efficient LLMs Inference Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:25:58.269821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T17:39:18.387881Z digest=sha256:24629099a4b810d910c35df72ab7adaf491a046a2eda7462abff7b1ad3705a5d