Pith. sign in

Paper Citation Record · LEDGER

Efficient Content-Based Sparse Attention with Routing Transformers

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2003.05997.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2003.05997 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T14:50:03.831572Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T19:45:36.414760Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e6bf9c7a-ee46-400a-93e2-c9a1bce4889b · inbound

Longformer: The Long-Document Transformer cites this paper.

Longformer: The Long-Document Transformer Efficient Content-Based Sparse Attention with Routing Transformers

Reference 115

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:29:58.812383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T13:29:58.719341Z digest=sha256:872c78c8b9d33904797aa68ddae845915336ad5a02703a62d5cb70efcf717c72

Observation 59d670d3-6313-40bb-a7d6-85ce5eef24f2 · inbound

Rethinking Attention with Performers cites this paper.

Rethinking Attention with Performers Efficient Content-Based Sparse Attention with Routing Transformers

Reference 147

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:16:14.467988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-12T09:16:14.336570Z digest=sha256:e2a98087721b1babbeac37b70db7cc3b0fc45a9df77bde09d8081d026ba95525

Observation b1bf7893-dbbe-421f-8b55-f8978425fd85 · inbound

Deformable DETR: Deformable Transformers for End-to-End Object Detection cites this paper.

Deformable DETR: Deformable Transformers for End-to-End Object Detection Efficient Content-Based Sparse Attention with Routing Transformers

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:47:17.010120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T09:47:16.915936Z digest=sha256:ed8908be78fb8bba7c7ce4ee840289eca58f1f80a93940edf014bbc56a146377

Observation 8747ea07-4842-4805-967b-f8a34b4a94f4 · inbound

PaLM: Scaling Language Modeling with Pathways cites this paper.

PaLM: Scaling Language Modeling with Pathways Efficient Content-Based Sparse Attention with Routing Transformers

Reference 129

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:45:07.356120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T23:45:06.755839Z digest=sha256:b6999e8e7a773efe0d57b45a7f6830bed82391f0a5598e4a11d44a0eba816cd1

Observation 16a64e48-e4bb-40c1-aaf5-bb23935fdda5 · inbound

Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads cites this paper.

Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads Efficient Content-Based Sparse Attention with Routing Transformers

Reference 241

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T10:36:18.322637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T10:36:17.764761Z digest=sha256:2eecd92b37d6b782a7611c0486e182d097ce0ad43bf3f9fb3f94fe55e230423b

Observation 27b91132-d50d-4583-87fa-e5e625d4ff0f · inbound

FlashAttention-3: Fast and Accurate Attention with Asynchrony and Low-precision cites this paper.

FlashAttention-3: Fast and Accurate Attention with Asynchrony and Low-precision Efficient Content-Based Sparse Attention with Routing Transformers

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:45:36.416769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T19:45:36.337956Z digest=sha256:62893ce15bff4f5ab1b997f3740cd7a3215a169abe63b665750bef3b598e827a

Observation 6b6f5401-5fb3-4b33-9ff4-d87bb92c3783 · inbound

Remembering Distinct Items, Not Tokens: A Learnable Dirichlet-Process Cache Between State-Space Models and Attention cites this paper.

Remembering Distinct Items, Not Tokens: A Learnable Dirichlet-Process Cache Between State-Space Models and Attention Efficient Content-Based Sparse Attention with Routing Transformers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-14T14:50:03.831572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T14:50:03.831572Z digest=sha256:8921327926b464babea4696acd1ded555c6b0900d956d62c1090ce3ecbd826be