Pith. sign in

Paper Citation Record · LEDGER

Modality Agnostic Efficient Long Range Encoder

As of 18 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 0 inbound Pith citation observations for arXiv:2507.19409.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.19409 v1

Coverage vector

measured 70 of 70 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:59:25.515340Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

70 of 70 outbound references displayed

  • verified exact2
  • verified fuzzy49
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40c324da-9dd3-4fb6-a6e1-339eec8d53bd · outbound

This paper cites GQA: Training gener- alized multi-query transformer models from multi-head checkpoints.

Modality Agnostic Efficient Long Range Encoder GQA: Training gener- alized multi-query transformer models from multi-head checkpoints

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.504863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.215137Z digest=sha256:d97ff3f7ed8947889786838f6c62eaab0171b94150d0ed2da229d5881cc02fcd

Observation d637edb3-9b63-4b0e-b8ee-fb77545f0f22 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Modality Agnostic Efficient Long Range Encoder Flamingo: a visual language model for few-shot learning

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.490859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.220029Z digest=sha256:722debbdbfd93597dcdb683afda542ce4595846e98b70b636b5441b135bb69cf

Observation 2ebd3bc5-ce57-4ab1-a20c-71e9b69b887a · outbound

This paper cites Longformer: The Long-Document Transformer.

Modality Agnostic Efficient Long Range Encoder Longformer: The Long-Document Transformer

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.224593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.224593Z digest=sha256:75e481ab01491a28f5e91a0b16d80ec4c0df046afd46b00f5ed683ea54516cbb

Observation 98a3b113-6bca-41bf-b8dd-b33dc0e8ff24 · outbound

This paper cites Token Merging: Your ViT But Faster.

Modality Agnostic Efficient Long Range Encoder Token Merging: Your ViT But Faster

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.229873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.229873Z digest=sha256:257f25250d0b21c3f8c0c9fe86fa1868387fc539b78760c4a9bfe6799ec0ea87

Observation 1383d3df-aadf-4b0f-8eb1-7afe67b8464e · outbound

This paper cites Scaling Transformer to 1M tokens and beyond with RMT.

Modality Agnostic Efficient Long Range Encoder Scaling Transformer to 1M tokens and beyond with RMT

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.234525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.234525Z digest=sha256:4b4395a212ddc30024cbd4b75a7c55297103130207fecd7c42b2a1e6dfc9a5e3

Observation e2ea8a0c-c9a0-4e14-ad68-0f5fd9b360b9 · outbound

This paper cites Vggsound: A large-scale audio-visual dataset.

Modality Agnostic Efficient Long Range Encoder Vggsound: A large-scale audio-visual dataset

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.476645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.239060Z digest=sha256:a58ba81a2f1097ad1049322aa73a5461f63cd38c393263f499e6d4a074fc48e1

Observation ce372dc8-eff8-473a-81cb-b13e67c1066e · outbound

This paper cites Classification of long sequential data using circular dilated convolutional neural networks.

Modality Agnostic Efficient Long Range Encoder Classification of long sequential data using circular dilated convolutional neural networks

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.462757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.244652Z digest=sha256:ab832f25c74dbf8b32d53ae47701a5e22b0a7a4509ae6acf19b11ab2fe594dba

Observation 3a3eded5-c146-4be6-b11e-5a5788db84ae · outbound

This paper cites Generating Long Sequences with Sparse Transformers.

Modality Agnostic Efficient Long Range Encoder Generating Long Sequences with Sparse Transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.248806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.248806Z digest=sha256:247c0eeb4bbcffb446179b2cec8d8a57dd00fbcfcbaeaf0cc81c492161845ef2

Observation 74a40dab-c92e-4313-884e-350194cfaae4 · outbound

This paper cites Rethinking attention with performers.

Modality Agnostic Efficient Long Range Encoder Rethinking attention with performers

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.449170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.253515Z digest=sha256:7844f7964aa75683b4545e26ecc90314a2da34ebf88cf43017857eab14091d89

Observation 6fc60df1-913f-441c-ac32-a11e3238eccb · outbound

This paper cites Ran- daugment: Practical automated data augmentation with a reduced search space.

Modality Agnostic Efficient Long Range Encoder Ran- daugment: Practical automated data augmentation with a reduced search space

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.435665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.258004Z digest=sha256:c71b76be4aef684cf5f9d07d5edec865ca311af9c92d6bd1b82651989f13f7d1

Observation 472b63c3-65ce-451d-b6bf-37fd00998efa · outbound

This paper cites an unresolved cited work.

Modality Agnostic Efficient Long Range Encoder Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:59:26.422263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.263129Z digest=sha256:6557b5cb09b9cad07dde462a66b429cb39889623a8e1f9dff3d512a771a9cb5b

Observation 17a97211-7920-42c0-b8bd-8b9d13558edf · outbound

This paper cites FlashAttention-2: Faster attention with better parallelism and work partitioning.

Modality Agnostic Efficient Long Range Encoder FlashAttention-2: Faster attention with better parallelism and work partitioning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.408224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.267331Z digest=sha256:6319343032bf84cf5dab4cb18bc2a2655b6f5e7942647864fcd89452834bb21a

Observation 7f27df0e-d055-413b-b62c-5d5f15659eab · outbound

This paper cites Fu, Stefano Ermon, Atri Rudra, and Christopher R´e.

Modality Agnostic Efficient Long Range Encoder Fu, Stefano Ermon, Atri Rudra, and Christopher R´e

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.394228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.271441Z digest=sha256:cca31f2969e3c94d85954234555a58ee35a915324ea548b800965a91a73616b7

Observation 1df443de-fc46-4f88-82e7-b5a1c91349ce · outbound

This paper cites The ucr time series classifica- tion archive, 2019.

Modality Agnostic Efficient Long Range Encoder The ucr time series classifica- tion archive, 2019

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.380830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.276528Z digest=sha256:1983dab51ec9232f0277a36999135b8d11093214517b2c25c645e2399fe92bac

Observation 70b88857-625a-4ccc-a378-f427f406048f · outbound

This paper cites DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model.

Modality Agnostic Efficient Long Range Encoder DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.280766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.280766Z digest=sha256:de586ccf8483d06c1c00844630dce31c001d32b44ea4ebf6db9a8f8d0572b899

Observation 9812f137-9952-4b9d-b4cb-9ff28c92b788 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Modality Agnostic Efficient Long Range Encoder Imagenet: A large-scale hierarchical image database

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.367324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.285702Z digest=sha256:3e6467a52e4da4aa4b1678631c9500d7e2d11c0b55c2a35a118b2704d8cfa2f0

Observation a36df706-28d2-411a-9e10-fe7894b98148 · outbound

This paper cites fvcore: A collection of core libraries for com- puter vision.

Modality Agnostic Efficient Long Range Encoder fvcore: A collection of core libraries for com- puter vision

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.353386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.290213Z digest=sha256:8fe935ed744c6205809ebf8ec9fe8a89552c32e0dc1973ec6bdfa7fc2a3e7391

Observation e9652e75-ee4d-4152-9a9b-e8763c39efae · outbound

This paper cites Multiscale vision transformers.

Modality Agnostic Efficient Long Range Encoder Multiscale vision transformers

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.339847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.294296Z digest=sha256:4fd3b73371abe20dfc4d147b05a2cf3dfe5c657c18c8f36ca73fe9434b2b99b1

Observation 2b5b81ca-432e-4d31-b44e-3f67d8965325 · outbound

This paper cites Hungry Hungry Hippos: Towards Language Modeling with State Space Models.

Modality Agnostic Efficient Long Range Encoder Hungry Hungry Hippos: Towards Language Modeling with State Space Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.298833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.298833Z digest=sha256:52ab17b626857e00cc725e218638fb50ed8ace7aeba0691acf424ae35c1ac6cd

Observation 76069de1-ba93-41d0-8855-dc3cca94f696 · outbound

This paper cites Dissecting Recall of Factual Associations in Auto-Regressive Language Models.

Modality Agnostic Efficient Long Range Encoder Dissecting Recall of Factual Associations in Auto-Regressive Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.303243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.303243Z digest=sha256:adfa5e8deb32832a8b21433d09c559e76849625537885e8c62a324aa9287a96f

Observation 6a1df863-2b17-47de-9a55-b342cc01a06b · outbound

This paper cites Liu, David Har- wath, Leonid Karlinsky, Hilde Kuehne, and James R.

Modality Agnostic Efficient Long Range Encoder Liu, David Har- wath, Leonid Karlinsky, Hilde Kuehne, and James R

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.326851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.307614Z digest=sha256:14b87a05b721865733d88fecb8f438cabf7894265ac08b9e5942c19596bd0912

Observation 832d6e1b-4a4b-44ba-9953-49dd3ec803f7 · outbound

This paper cites Choudhury, Saurabh M.

Modality Agnostic Efficient Long Range Encoder Choudhury, Saurabh M

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.313722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.311632Z digest=sha256:6d4d79e4096553181b0c633d52bfe6d531827a61ce1042717d996ab575e58b8f

Observation eca2bdb5-abcb-42e0-b621-7cea2e99dad0 · outbound

This paper cites Levit: a vision transformer in convnet’s clothing for faster inference.

Modality Agnostic Efficient Long Range Encoder Levit: a vision transformer in convnet’s clothing for faster inference

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.300518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.315688Z digest=sha256:1b49114ac53c4be0b0756090d5ca1901a112347babf9a2876b6716e34ff93eca

Observation 3598c43a-b2cc-4556-a459-144cd849f4f0 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Modality Agnostic Efficient Long Range Encoder Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.319687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.319687Z digest=sha256:e2de0cb179c2ced8620ead0ee822a03187f8cb142cf85b24ca100ba975a722be

Observation f192dcbc-ca27-40f6-8fa0-9b4667330089 · outbound

This paper cites Longt5: Efficient text-to-text transformer for long sequences.

Modality Agnostic Efficient Long Range Encoder Longt5: Efficient text-to-text transformer for long sequences

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.287274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.323970Z digest=sha256:4010c9dd82fe92660ccc42bca079ef09686b00ed49083798854a25e61d74081c

Observation b4763ef0-e95b-4a4e-b44d-067276310a14 · outbound

This paper cites Flatten transformer: Vision transformer using focused linear atten- tion.

Modality Agnostic Efficient Long Range Encoder Flatten transformer: Vision transformer using focused linear atten- tion

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.273407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.327908Z digest=sha256:397f21b46d94b3e6ff9d9dde8b30ff829360d81d06f019226c5fd5690fa818a9

Observation 6709e51d-7c88-4e42-be09-80c2fc9eafba · outbound

This paper cites Masked autoencoders are scalable vision learn- ers.

Modality Agnostic Efficient Long Range Encoder Masked autoencoders are scalable vision learn- ers

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.260553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.332340Z digest=sha256:7162b5aaf6fd7e827264a4d3832d9ca04a69f33b33fe89eadf978c9db5956187

Observation f215666b-23e2-43db-b069-4e43ac9dcf0f · outbound

This paper cites Rethinking spatial dimensions of vision transformers.

Modality Agnostic Efficient Long Range Encoder Rethinking spatial dimensions of vision transformers

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.247518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.336473Z digest=sha256:cff95e06608978fd9d160d0623e9fff6eaa9fca79ecb2e8170a3040e10858e95

Observation a773974a-63ca-4161-be25-957645d5e75d · outbound

This paper cites an unresolved cited work.

Modality Agnostic Efficient Long Range Encoder Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:59:26.233896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.340570Z digest=sha256:0ca183a4502ad857bf77f83ea6b8b5293f825e313b4acb09784a1d3706716c2c

Observation 38e4425a-a289-466f-888a-9a71f47c3090 · outbound

This paper cites Deep networks with stochastic depth.

Modality Agnostic Efficient Long Range Encoder Deep networks with stochastic depth

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.219846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.344819Z digest=sha256:976808182cd291d66e5cc67375689742839897f7ebc1053cdf55c47d246a6d75

Observation b35ca5c3-3979-435c-8e28-49b8e3f78cfc · outbound

This paper cites Transformers: State-of-the-art machine learning for pytorch, tensorflow, and jax.

Modality Agnostic Efficient Long Range Encoder Transformers: State-of-the-art machine learning for pytorch, tensorflow, and jax

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.205532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.348878Z digest=sha256:d9636197a9eea5e29a571d3a65a18d642804e1b015e8f79b2efe474c60ef4d23

Observation a8d10155-e7e8-4edf-b305-c51f4b269ddc · outbound

This paper cites An empirical survey of data augmentation for time series classification with neural networks.

Modality Agnostic Efficient Long Range Encoder An empirical survey of data augmentation for time series classification with neural networks

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.191554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.352830Z digest=sha256:e765aaa4ce411cb43919ade7f79e88e4bd76fe8620dc0943730fce7987f02a94

Observation 2e913e1d-44f5-4c50-98d0-0ec554898e92 · outbound

This paper cites Katharopoulos, A.

Modality Agnostic Efficient Long Range Encoder Katharopoulos, A

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.177627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.357165Z digest=sha256:633e55e7e7392fd76f47f8b9f4b29248ad528f5d72cb0095892bc68e36752b2f

Observation 09c1b765-c89c-42dc-8aba-31811d20d31c · outbound

This paper cites Reformer: The efficient transformer.

Modality Agnostic Efficient Long Range Encoder Reformer: The efficient transformer

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.164371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.361347Z digest=sha256:a191ce5efda602366aeecdfda252b8cbe896a4b1d81573bc6af04c80d9ed44e1

Observation f2a17ed6-74b4-4c32-97ab-2dc84f820da3 · outbound

This paper cites Sequence parallelism: Long sequence training from sys- tem perspective.

Modality Agnostic Efficient Long Range Encoder Sequence parallelism: Long sequence training from sys- tem perspective

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.151210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.365969Z digest=sha256:64e1aea76c865940f362f9d347b6f55a33d1c36f30d66ed646729ff02282a4c5

Observation 8b3e01ac-e52f-415a-a6f3-2454d57caa58 · outbound

This paper cites Mvitv2: Im- proved multiscale vision transformers for classification and detec- tion.

Modality Agnostic Efficient Long Range Encoder Mvitv2: Im- proved multiscale vision transformers for classification and detec- tion

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.137162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.370122Z digest=sha256:8711cda510b993750737fe38724359bc01b7d08af0aa17217716be79f1994b38

Observation bb17e37d-ac4b-43f8-8176-fa7524e45327 · outbound

This paper cites Efficientformer: Vi- sion transformers at mobilenet speed.

Modality Agnostic Efficient Long Range Encoder Efficientformer: Vi- sion transformers at mobilenet speed

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.123873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.374058Z digest=sha256:bb36d55ea76e4f1af0d2a533021e3c0b2d008008db50255c9d02ff457b7fe1e2

Observation f6ee13b4-36fc-43fa-b932-88a4db6f26c0 · outbound

This paper cites Ring Attention with Blockwise Transformers for Near-Infinite Context.

Modality Agnostic Efficient Long Range Encoder Ring Attention with Blockwise Transformers for Near-Infinite Context

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.378070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.378070Z digest=sha256:a93d5a0378c629079bd4d8600ad852543ac2981ca837249f39f0ced7504b268a

Observation 7b8adc2b-ecdf-4c52-8683-59e0fadc48ae · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

Modality Agnostic Efficient Long Range Encoder RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.382444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.382444Z digest=sha256:33645fe95af335587739a577f6c3b0e0baa41c6de2f4df3f135b32dcfc9fd9f8

Observation 918ca2aa-219e-424f-8732-02491fe6a18d · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

Modality Agnostic Efficient Long Range Encoder Swin transformer: Hierarchical vision transformer using shifted windows

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.111004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.386674Z digest=sha256:b87252ab4c0d08a05b07bdf32cea6e895bf16b6207581701b28999f2e83ea1c0

Observation 23d30d51-e82c-401f-9cf4-4f67d03a0e5d · outbound

This paper cites Sgdr: Stochastic gradient descent with warm restarts.

Modality Agnostic Efficient Long Range Encoder Sgdr: Stochastic gradient descent with warm restarts

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.097838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.390792Z digest=sha256:7656ef231a846f42436033bd83b8c72cbaefe4492eb0160442e98c08a6332ce9

Observation afaefcc2-3ff1-45a9-85a1-e3a223656ba1 · outbound

This paper cites Decoupled weight decay regular- ization.

Modality Agnostic Efficient Long Range Encoder Decoupled weight decay regular- ization

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.083666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.394907Z digest=sha256:70b14dface536d489fb4e7971ffda864b6844d46331c6acca30270b1c74d0927

Observation d6781f67-41c5-4bf6-b199-d4e5fac58704 · outbound

This paper cites Token pooling in vi- sion transformers for image classification.

Modality Agnostic Efficient Long Range Encoder Token pooling in vi- sion transformers for image classification

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.069938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.399303Z digest=sha256:dec96462d49e5c71ec7daf7b419b7f080bec11b547204e32c66fa941433795b7

Observation ded4bb0b-b8c5-43c3-ad37-3e8c7564c964 · outbound

This paper cites Lo- cating and editing factual associations in gpt.

Modality Agnostic Efficient Long Range Encoder Lo- cating and editing factual associations in gpt

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.055908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.403328Z digest=sha256:6832aeab03802e36326f09172675e4503d5f3822f2480dfab28e5950d1a077a8

Observation de3fd62f-775f-42dc-bcbe-5e061e48957c · outbound

This paper cites Scalable vision transformers with hierarchical pooling.

Modality Agnostic Efficient Long Range Encoder Scalable vision transformers with hierarchical pooling

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.042794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.407321Z digest=sha256:4b71ada553f63e67b53c79c16ed10af2fc31e21b3237dcd31a1485fa00b15d87

Observation e5223b9e-3647-40da-982b-cb780d597be6 · outbound

This paper cites Investigating efficiently ex- tending transformers for long input summarization.

Modality Agnostic Efficient Long Range Encoder Investigating efficiently ex- tending transformers for long input summarization

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.029374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.411439Z digest=sha256:5641b72022c9adcfd83a3f645e6971c08aaa0147b2574130459ce2b50b3efa95

Observation 43ac81ea-8578-4a32-91fe-38ac259db8fc · outbound

This paper cites Self-attention Does Not Need $O(n^2)$ Memory.

Modality Agnostic Efficient Long Range Encoder Self-attention Does Not Need $O(n^2)$ Memory

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.415528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.415528Z digest=sha256:1e469d4c2eaf5195c364d3878828b673d9a6054def39afbf1f2882340e820bb3

Observation ea68c8ca-08bd-443e-9ac4-bcea0df233a0 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

Modality Agnostic Efficient Long Range Encoder Learning Transferable Visual Models From Natural Language Supervision

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.419888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.419888Z digest=sha256:328bdf8129a983f3d660597497e8ae0a5c7698af45a8b8712c7e25065a12b663

Observation cf490f3b-ba2d-4c84-bc08-b07d41e60581 · outbound

This paper cites Fast Transformer Decoding: One Write-Head is All You Need.

Modality Agnostic Efficient Long Range Encoder Fast Transformer Decoding: One Write-Head is All You Need

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.424102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.424102Z digest=sha256:aad6d6bd919f24f838eb20add84471278754438ac6d72fb8faeb1fa089f71126

Observation 8e568b0c-f1df-468c-a9e3-1b02f0a5ae8c · outbound

This paper cites Dropout: A simple way to prevent neural networks from overfitting.

Modality Agnostic Efficient Long Range Encoder Dropout: A simple way to prevent neural networks from overfitting

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.015343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.428525Z digest=sha256:d2abd88723b7dadf8ea69b63249d008b4cc4c4ca4e7dbf94edd62f062c696dbf

Observation d6b1d789-7f89-4778-bc55-3f9a920933e9 · outbound

This paper cites Scaling Granite Code Models to 128K Context.

Modality Agnostic Efficient Long Range Encoder Scaling Granite Code Models to 128K Context

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.433055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.433055Z digest=sha256:a849b1161687ff5c84463fcbf243d1f6c6f1a854468d883479103eee9e90dda2

Observation 8b34aa8f-5245-4abd-bf37-e9ac6d438833 · outbound

This paper cites Rethinking the inception architecture for computer vision.

Modality Agnostic Efficient Long Range Encoder Rethinking the inception architecture for computer vision

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.002442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.438136Z digest=sha256:87d90dae66f3c93b9a6444f794482c15bfba48f5bce7c11dd30767bb0be1517c

Observation 27c25218-4a4c-4e58-b892-c55857594733 · outbound

This paper cites Sparse sinkhorn attention.

Modality Agnostic Efficient Long Range Encoder Sparse sinkhorn attention

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.988808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.442255Z digest=sha256:05eb111dc697a956254bc3b2d139c4903e49f8010ee8bf4e542e893fd6a1e1ee

Observation c568d9a2-a81d-4380-97b2-11a0af045455 · outbound

This paper cites Long range arena: A benchmark for efficient transformers.

Modality Agnostic Efficient Long Range Encoder Long range arena: A benchmark for efficient transformers

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.975646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.446268Z digest=sha256:049fab62a0eeee1c8d1e9959ee3a1d1231534c281d4420e73c8675dba6ecf9a3

Observation 9b3e551d-34f2-4704-b71c-0c858959ce64 · outbound

This paper cites Atten- tion is all you need.

Modality Agnostic Efficient Long Range Encoder Atten- tion is all you need

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.962476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.450471Z digest=sha256:56ac0caa87f19850b024f2bcd0e8b09681859bba5e5857883df30d1b04289095

Observation 27f9710a-aefa-4135-8eab-6c0349fa9bd2 · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

Modality Agnostic Efficient Long Range Encoder Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.454731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.454731Z digest=sha256:c37583f8350dd076ef1c19b949c11676eb142d5208ada37b04076e33f00693f5

Observation d3cefd6b-9bfa-48fd-bdbe-9f1891eb5021 · outbound

This paper cites Huang, Krzysztof Choromanski, Valerii Likhosherstov, and Adrian Weller.

Modality Agnostic Efficient Long Range Encoder Huang, Krzysztof Choromanski, Valerii Likhosherstov, and Adrian Weller

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.948665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.459274Z digest=sha256:8aa4fcd11a35c37c4332f281c78149d2349561e15f103e24eee2e8b9ca863ec6

Observation 205864dc-13d0-4511-9f0b-658740b3868d · outbound

This paper cites MeSHup: A Corpus for Full Text Biomedical Document Indexing.

Modality Agnostic Efficient Long Range Encoder MeSHup: A Corpus for Full Text Biomedical Document Indexing

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:59:25.605672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.464664Z digest=sha256:1d461006a512a92be5a3f47d65b6fba63b59c2c388f929b42c09fa58a78922b0

Observation aed7673b-bfe3-4ea9-8b30-3d74e96e3ed5 · outbound

This paper cites Pytorch image models.

Modality Agnostic Efficient Long Range Encoder Pytorch image models

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.934518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.469293Z digest=sha256:65edcceba8a77576394b66aabb37354c9ad55746a11124d9c1a2d1b8d0d061d1

Observation e5fe8155-0781-420d-a003-e7eaaacebf1c · outbound

This paper cites Ef- fective long-context scaling of foundation models.

Modality Agnostic Efficient Long Range Encoder Ef- fective long-context scaling of foundation models

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.920457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.473274Z digest=sha256:bdc8127cf8da035f4f8f70643372da735e41b57e7fe9957af816b48ffec678f2

Observation 1da600db-4868-46d7-ac0d-f7ffa5f80bf4 · outbound

This paper cites A-vit: Adaptive tokens for efficient vision transformer.

Modality Agnostic Efficient Long Range Encoder A-vit: Adaptive tokens for efficient vision transformer

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.906773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.477372Z digest=sha256:c251fbcc83364a4cc7336cd9cb234a495a43ba7dc20c1235ad3eb3593fb89d13

Observation 65337461-5d63-4e10-b756-0c38205a0dbe · outbound

This paper cites MEGABYTE: Predicting Million-byte Sequences with Multiscale Transformers.

Modality Agnostic Efficient Long Range Encoder MEGABYTE: Predicting Million-byte Sequences with Multiscale Transformers

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.481718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.481718Z digest=sha256:4f8908de801507dd708cf5bdfff3584a926bb3a11197a3d670ed096fc1a682b5

Observation b1a65efb-5ebd-44f3-bf4f-7a6b80d8256e · outbound

This paper cites Metaformer is actually what you need for vision.

Modality Agnostic Efficient Long Range Encoder Metaformer is actually what you need for vision

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.892725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.486110Z digest=sha256:47263e93986ae8c829bb0e2517d259518dca94f87702d627a9230a6cce3c603f

Observation c4b9783c-bed3-400d-8f7c-dd4510e5d740 · outbound

This paper cites Cutmix: Regularization strategy to train strong classifiers with localizable features.

Modality Agnostic Efficient Long Range Encoder Cutmix: Regularization strategy to train strong classifiers with localizable features

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.877463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.490295Z digest=sha256:148fb7c2d26009b6198b66774d90dd49754b8d148d33b0c884b210373f8ed56a

Observation 1768df0d-4863-4051-8548-0f4155c2aea5 · outbound

This paper cites Big bird: Trans- formers for longer sequences.

Modality Agnostic Efficient Long Range Encoder Big bird: Trans- formers for longer sequences

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.863098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.494565Z digest=sha256:c4924166a70f2f9c82107703f00521ce7fac3a529f1679bf4d6dc31dddb8d489

Observation c6651615-4d64-4f29-b1cc-0edb604dc96f · outbound

This paper cites mixup: Beyond empirical risk minimization.

Modality Agnostic Efficient Long Range Encoder mixup: Beyond empirical risk minimization

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.849604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.498610Z digest=sha256:9b8d643e92b8446dd6de30b661528624cf8f337e39fc35b3d14e8fde1e800cc0

Observation 39cd3b71-1fcd-483c-8f13-9eaaa2c1e444 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

Modality Agnostic Efficient Long Range Encoder OPT: Open Pre-trained Transformer Language Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.502868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.502868Z digest=sha256:a0af2008a745845b5715179b845ebfa9b95f7e685dc8e1d5946161194ee40791

Observation 9db88336-7389-4943-966a-6d4c20ab901c · outbound

This paper cites Adavit: Adaptive vision transformers for efficient image recognition.

Modality Agnostic Efficient Long Range Encoder Adavit: Adaptive vision transformers for efficient image recognition

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.835846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.507154Z digest=sha256:932322b778be67626781089fb7d512811de7a5fece482caab137108c7e338dc3

Observation 27b8126e-4ca5-4088-86e3-ad1eaa430b5f · outbound

This paper cites Long-short transformer: Efficient transformers for language and vision.

Modality Agnostic Efficient Long Range Encoder Long-short transformer: Efficient transformers for language and vision

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.822039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.511224Z digest=sha256:9cfce394fca5059b3cdf93669274d1c2318fcf611418e88361b30dd2c7eddea9

Observation feabdaa6-8dd1-42ba-857f-21e471ac9c98 · outbound

This paper cites Multiscale Audio Spectrogram Transformer for Efficient Audio Classification.

Modality Agnostic Efficient Long Range Encoder Multiscale Audio Spectrogram Transformer for Efficient Audio Classification

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:59:25.556777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T17:59:25.515340Z digest=sha256:73c5aef2a2de98c68dfd5b3fe322068b9f85038b7db85164be57d2b1b2495bcc

Pith citing papers

No inbound Pith citation observations are available.