Pith. sign in

Paper Citation Record · LEDGER

Modality Agnostic Efficient Long Range Encoder

As of 22 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 0 inbound Pith citation observations for arXiv:2507.19409.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.19409 v1

Coverage vector

measured 70 of 70 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:59:25.515340Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

70 of 70 outbound references displayed

  • verified exact2
  • verified fuzzy49
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40c324da-9dd3-4fb6-a6e1-339eec8d53bd · outbound

This paper cites GQA: Training gener- alized multi-query transformer models from multi-head checkpoints.

Modality Agnostic Efficient Long Range Encoder GQA: Training gener- alized multi-query transformer models from multi-head checkpoints

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.504863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.215137Z digest=sha256:a17771c4c5d2976ca99fd7363fdc188bae6f13a8a7e848aab01742ad51ffb42f

Observation d637edb3-9b63-4b0e-b8ee-fb77545f0f22 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Modality Agnostic Efficient Long Range Encoder Flamingo: a visual language model for few-shot learning

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.490859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.220029Z digest=sha256:d41c0ace71c4c38a0c9b87d8c8f57801c92ca6d0c256095c860fc2aa414955c7

Observation 2ebd3bc5-ce57-4ab1-a20c-71e9b69b887a · outbound

This paper cites Longformer: The Long-Document Transformer.

Modality Agnostic Efficient Long Range Encoder Longformer: The Long-Document Transformer

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.224593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.224593Z digest=sha256:d5e0d3a8d6718d1d47fdce930dc45cbc20e745f3b4722d10840d55c034f84db2

Observation 98a3b113-6bca-41bf-b8dd-b33dc0e8ff24 · outbound

This paper cites Token Merging: Your ViT But Faster.

Modality Agnostic Efficient Long Range Encoder Token Merging: Your ViT But Faster

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.229873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.229873Z digest=sha256:ed7368ca51df4fdfb7ac216067b459dd41e3eae6e450306b90d826b5b7f9b5de

Observation 1383d3df-aadf-4b0f-8eb1-7afe67b8464e · outbound

This paper cites Scaling Transformer to 1M tokens and beyond with RMT.

Modality Agnostic Efficient Long Range Encoder Scaling Transformer to 1M tokens and beyond with RMT

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.234525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.234525Z digest=sha256:990c56536c4f34a0f4bf413cb9870a7849172202227e8282f72fcc777c320231

Observation e2ea8a0c-c9a0-4e14-ad68-0f5fd9b360b9 · outbound

This paper cites Vggsound: A large-scale audio-visual dataset.

Modality Agnostic Efficient Long Range Encoder Vggsound: A large-scale audio-visual dataset

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.476645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.239060Z digest=sha256:3744a1263538dafa41c7c3484ee12d1ec997aacbe6021a83a54b13056f862a41

Observation ce372dc8-eff8-473a-81cb-b13e67c1066e · outbound

This paper cites Classification of long sequential data using circular dilated convolutional neural networks.

Modality Agnostic Efficient Long Range Encoder Classification of long sequential data using circular dilated convolutional neural networks

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.462757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.244652Z digest=sha256:caa9b2e3dc968f3bfa58bf26679157ef60747e5d8a8e0bdcc7276be54a51eec4

Observation 3a3eded5-c146-4be6-b11e-5a5788db84ae · outbound

This paper cites Generating Long Sequences with Sparse Transformers.

Modality Agnostic Efficient Long Range Encoder Generating Long Sequences with Sparse Transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.248806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.248806Z digest=sha256:295580ee25509af754e2f996e8352688669064fbfdf576150bcc1917804ad2e6

Observation 74a40dab-c92e-4313-884e-350194cfaae4 · outbound

This paper cites Rethinking attention with performers.

Modality Agnostic Efficient Long Range Encoder Rethinking attention with performers

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.449170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.253515Z digest=sha256:a47aedd0d1aa47feb3fa94ea4f5538ea07a84b1714a1689040c7c94165d224fa

Observation 6fc60df1-913f-441c-ac32-a11e3238eccb · outbound

This paper cites Ran- daugment: Practical automated data augmentation with a reduced search space.

Modality Agnostic Efficient Long Range Encoder Ran- daugment: Practical automated data augmentation with a reduced search space

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.435665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.258004Z digest=sha256:0a52cf09db463168f5ff337fbab569b18dd3c5756ee9d5b0f5933b96dd487643

Observation 472b63c3-65ce-451d-b6bf-37fd00998efa · outbound

This paper cites an unresolved cited work.

Modality Agnostic Efficient Long Range Encoder Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:59:26.422263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.263129Z digest=sha256:83998ea21331170d7425a6f341393ced481ddef897f136e0c83cdc15a8d7ebf5

Observation 17a97211-7920-42c0-b8bd-8b9d13558edf · outbound

This paper cites FlashAttention-2: Faster attention with better parallelism and work partitioning.

Modality Agnostic Efficient Long Range Encoder FlashAttention-2: Faster attention with better parallelism and work partitioning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.408224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.267331Z digest=sha256:87f5214809729b0a77fce408757ed334611d51823d25653c238c750e322f45b9

Observation 7f27df0e-d055-413b-b62c-5d5f15659eab · outbound

This paper cites Fu, Stefano Ermon, Atri Rudra, and Christopher R´e.

Modality Agnostic Efficient Long Range Encoder Fu, Stefano Ermon, Atri Rudra, and Christopher R´e

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.394228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.271441Z digest=sha256:5e8780ed42711dc4bce299a5633b976efc1187b99b2ff9e442420ca05628ed7c

Observation 1df443de-fc46-4f88-82e7-b5a1c91349ce · outbound

This paper cites The ucr time series classifica- tion archive, 2019.

Modality Agnostic Efficient Long Range Encoder The ucr time series classifica- tion archive, 2019

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.380830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.276528Z digest=sha256:25b98da1a253dd9297fde7314e2f26a05dac9c33d8bbf7b5a634f069779faf3e

Observation 70b88857-625a-4ccc-a378-f427f406048f · outbound

This paper cites DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model.

Modality Agnostic Efficient Long Range Encoder DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.280766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.280766Z digest=sha256:4d79852f2ad20e90a2f92f3c4ec845f09f79899bd587920a221ecf16b3e3e429

Observation 9812f137-9952-4b9d-b4cb-9ff28c92b788 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Modality Agnostic Efficient Long Range Encoder Imagenet: A large-scale hierarchical image database

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.367324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.285702Z digest=sha256:e0340e2bd3a1969b49bc92ea0a11e9250ba667afb39be769d90ea328bca9cae2

Observation a36df706-28d2-411a-9e10-fe7894b98148 · outbound

This paper cites fvcore: A collection of core libraries for com- puter vision.

Modality Agnostic Efficient Long Range Encoder fvcore: A collection of core libraries for com- puter vision

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.353386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.290213Z digest=sha256:df623a48c3e33c3aaac167820d38057e27918d643fc65d2723dd92b74f06c1ab

Observation e9652e75-ee4d-4152-9a9b-e8763c39efae · outbound

This paper cites Multiscale vision transformers.

Modality Agnostic Efficient Long Range Encoder Multiscale vision transformers

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.339847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.294296Z digest=sha256:335033f840c53ac207a8c73e81871ce3338347b308dc7b2de507ea936b99cfa8

Observation 2b5b81ca-432e-4d31-b44e-3f67d8965325 · outbound

This paper cites Hungry Hungry Hippos: Towards Language Modeling with State Space Models.

Modality Agnostic Efficient Long Range Encoder Hungry Hungry Hippos: Towards Language Modeling with State Space Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.298833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.298833Z digest=sha256:0bac952fe66d5f28f57b70e8cb4dfbe3725276d394e35c26844d6035560143d8

Observation 76069de1-ba93-41d0-8855-dc3cca94f696 · outbound

This paper cites Dissecting Recall of Factual Associations in Auto-Regressive Language Models.

Modality Agnostic Efficient Long Range Encoder Dissecting Recall of Factual Associations in Auto-Regressive Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.303243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.303243Z digest=sha256:40ca76d3fa064e8d289d530e6595b6cd504320dc9e11736742375ef6ad757107

Observation 6a1df863-2b17-47de-9a55-b342cc01a06b · outbound

This paper cites Liu, David Har- wath, Leonid Karlinsky, Hilde Kuehne, and James R.

Modality Agnostic Efficient Long Range Encoder Liu, David Har- wath, Leonid Karlinsky, Hilde Kuehne, and James R

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.326851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.307614Z digest=sha256:b8a405610625ad889b2bc5f8f404a748c181a76c96429730a362ac2d143a1930

Observation 832d6e1b-4a4b-44ba-9953-49dd3ec803f7 · outbound

This paper cites Choudhury, Saurabh M.

Modality Agnostic Efficient Long Range Encoder Choudhury, Saurabh M

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.313722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.311632Z digest=sha256:861e76040ac884866e798fac088e7b9832283d75e408a8578dcfe4353d563028

Observation eca2bdb5-abcb-42e0-b621-7cea2e99dad0 · outbound

This paper cites Levit: a vision transformer in convnet’s clothing for faster inference.

Modality Agnostic Efficient Long Range Encoder Levit: a vision transformer in convnet’s clothing for faster inference

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.300518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.315688Z digest=sha256:81e23d108f1710c436cfefd0a2dc25cc23dcb40a506c9b9e38754c14a08e9d79

Observation 3598c43a-b2cc-4556-a459-144cd849f4f0 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Modality Agnostic Efficient Long Range Encoder Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.319687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.319687Z digest=sha256:b6b754c8c0daa0994d613d7cab45eb6963bb380ca13aef964ee8e9f6b5400952

Observation f192dcbc-ca27-40f6-8fa0-9b4667330089 · outbound

This paper cites Longt5: Efficient text-to-text transformer for long sequences.

Modality Agnostic Efficient Long Range Encoder Longt5: Efficient text-to-text transformer for long sequences

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.287274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.323970Z digest=sha256:29ad8b2f690a2a6c83ff8807717307ed33a3c9a0ed2604bfc8c4ae3b489f8f73

Observation b4763ef0-e95b-4a4e-b44d-067276310a14 · outbound

This paper cites Flatten transformer: Vision transformer using focused linear atten- tion.

Modality Agnostic Efficient Long Range Encoder Flatten transformer: Vision transformer using focused linear atten- tion

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.273407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.327908Z digest=sha256:818304c0eeeb037d54d49893bec706a127302aa5acb8bd8c163025eb81b88011

Observation 6709e51d-7c88-4e42-be09-80c2fc9eafba · outbound

This paper cites Masked autoencoders are scalable vision learn- ers.

Modality Agnostic Efficient Long Range Encoder Masked autoencoders are scalable vision learn- ers

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.260553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.332340Z digest=sha256:ecf1ed4325c922af3efea80e94de60b01afeb94f689e15de719e194d2a5a3a27

Observation f215666b-23e2-43db-b069-4e43ac9dcf0f · outbound

This paper cites Rethinking spatial dimensions of vision transformers.

Modality Agnostic Efficient Long Range Encoder Rethinking spatial dimensions of vision transformers

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.247518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.336473Z digest=sha256:19227282f6c7d4f09ba761d42547fc3c9edd2346f45146a239eec918a7547e83

Observation a773974a-63ca-4161-be25-957645d5e75d · outbound

This paper cites an unresolved cited work.

Modality Agnostic Efficient Long Range Encoder Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T17:59:26.233896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.340570Z digest=sha256:8791d6a71320dcbcfd670d5c02ef3ca6f8d73c483bc04b333be739ca090fb70f

Observation 38e4425a-a289-466f-888a-9a71f47c3090 · outbound

This paper cites Deep networks with stochastic depth.

Modality Agnostic Efficient Long Range Encoder Deep networks with stochastic depth

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.219846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.344819Z digest=sha256:395ed5841a119a9907fbc60dda20af0ab145217259738b90e332ec8685adb8e7

Observation b35ca5c3-3979-435c-8e28-49b8e3f78cfc · outbound

This paper cites Transformers: State-of-the-art machine learning for pytorch, tensorflow, and jax.

Modality Agnostic Efficient Long Range Encoder Transformers: State-of-the-art machine learning for pytorch, tensorflow, and jax

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.205532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.348878Z digest=sha256:09a360280003eebabc8f252cbc846e58346f042750fa586d5842a4d2f0a04d3a

Observation a8d10155-e7e8-4edf-b305-c51f4b269ddc · outbound

This paper cites An empirical survey of data augmentation for time series classification with neural networks.

Modality Agnostic Efficient Long Range Encoder An empirical survey of data augmentation for time series classification with neural networks

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.191554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.352830Z digest=sha256:4502e7955e1feccb48ead1f73d00b4a836c8d90db2c9a6ec5fd75cd271d77311

Observation 2e913e1d-44f5-4c50-98d0-0ec554898e92 · outbound

This paper cites Katharopoulos, A.

Modality Agnostic Efficient Long Range Encoder Katharopoulos, A

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.177627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.357165Z digest=sha256:100faa31978055bd67fb0c57a38fcbd4b839e32930bb70eca974fef82fe4ff66

Observation 09c1b765-c89c-42dc-8aba-31811d20d31c · outbound

This paper cites Reformer: The efficient transformer.

Modality Agnostic Efficient Long Range Encoder Reformer: The efficient transformer

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.164371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.361347Z digest=sha256:8f2d072b8038357c8e46ac0099d983ccfc98089c80253e599e98631b32d98766

Observation f2a17ed6-74b4-4c32-97ab-2dc84f820da3 · outbound

This paper cites Sequence parallelism: Long sequence training from sys- tem perspective.

Modality Agnostic Efficient Long Range Encoder Sequence parallelism: Long sequence training from sys- tem perspective

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.151210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.365969Z digest=sha256:e5efe087d459569b52686b2c72189e59fbf465b136334c4f78919ada5ffb6a81

Observation 8b3e01ac-e52f-415a-a6f3-2454d57caa58 · outbound

This paper cites Mvitv2: Im- proved multiscale vision transformers for classification and detec- tion.

Modality Agnostic Efficient Long Range Encoder Mvitv2: Im- proved multiscale vision transformers for classification and detec- tion

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.137162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.370122Z digest=sha256:f3c77b144a6138fae596618c7af1dc4f15d73d2dbb8a4cc1bc66e802b7b89dea

Observation bb17e37d-ac4b-43f8-8176-fa7524e45327 · outbound

This paper cites Efficientformer: Vi- sion transformers at mobilenet speed.

Modality Agnostic Efficient Long Range Encoder Efficientformer: Vi- sion transformers at mobilenet speed

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.123873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.374058Z digest=sha256:a3c9a141dc42d6af1a17e735a6ad4802bbbcd2e4f0ccecdf9a8cea41b731bb09

Observation f6ee13b4-36fc-43fa-b932-88a4db6f26c0 · outbound

This paper cites Ring Attention with Blockwise Transformers for Near-Infinite Context.

Modality Agnostic Efficient Long Range Encoder Ring Attention with Blockwise Transformers for Near-Infinite Context

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.378070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.378070Z digest=sha256:f74c3e592bd6fee0b0b9e3071d079d489b2005abcc32fe670b19e0d1b61f0cca

Observation 7b8adc2b-ecdf-4c52-8683-59e0fadc48ae · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

Modality Agnostic Efficient Long Range Encoder RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.382444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.382444Z digest=sha256:043c63e5a3032ecbf6642168799825c8d20c45b5c4882df0a5f5ecebd7c87d2d

Observation 918ca2aa-219e-424f-8732-02491fe6a18d · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

Modality Agnostic Efficient Long Range Encoder Swin transformer: Hierarchical vision transformer using shifted windows

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.111004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.386674Z digest=sha256:4bf6ecfe1f71e4251432d72a4d0d356aa23d0175f6a832c7906bef977f2878be

Observation 23d30d51-e82c-401f-9cf4-4f67d03a0e5d · outbound

This paper cites Sgdr: Stochastic gradient descent with warm restarts.

Modality Agnostic Efficient Long Range Encoder Sgdr: Stochastic gradient descent with warm restarts

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.097838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.390792Z digest=sha256:58c17329d23068987cff67ae6908ea12185f6b6216e52d8b4cb18301999b9449

Observation afaefcc2-3ff1-45a9-85a1-e3a223656ba1 · outbound

This paper cites Decoupled weight decay regular- ization.

Modality Agnostic Efficient Long Range Encoder Decoupled weight decay regular- ization

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.083666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.394907Z digest=sha256:cc9cedde18237e62a26525bbd34d0baf3b8d9556af623929fcef80817902ec12

Observation d6781f67-41c5-4bf6-b199-d4e5fac58704 · outbound

This paper cites Token pooling in vi- sion transformers for image classification.

Modality Agnostic Efficient Long Range Encoder Token pooling in vi- sion transformers for image classification

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.069938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.399303Z digest=sha256:42a6f3fc36cba6bf6191f1b300be0ec789d6ffc3713a87088780d3067955991b

Observation ded4bb0b-b8c5-43c3-ad37-3e8c7564c964 · outbound

This paper cites Lo- cating and editing factual associations in gpt.

Modality Agnostic Efficient Long Range Encoder Lo- cating and editing factual associations in gpt

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.055908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.403328Z digest=sha256:abe4f9e84dbd62da0898756edce11d728b45d3709c2a33683959844bcf3def57

Observation de3fd62f-775f-42dc-bcbe-5e061e48957c · outbound

This paper cites Scalable vision transformers with hierarchical pooling.

Modality Agnostic Efficient Long Range Encoder Scalable vision transformers with hierarchical pooling

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.042794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.407321Z digest=sha256:02cb21766e130831ccbdad2448ae514bcf6e250a209caab1cfc0ac32d096ccb9

Observation e5223b9e-3647-40da-982b-cb780d597be6 · outbound

This paper cites Investigating efficiently ex- tending transformers for long input summarization.

Modality Agnostic Efficient Long Range Encoder Investigating efficiently ex- tending transformers for long input summarization

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.029374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.411439Z digest=sha256:f8de2e1e6847d5090e16b76e7bf097eb07698759cf77749eb92ccf32374433a3

Observation 43ac81ea-8578-4a32-91fe-38ac259db8fc · outbound

This paper cites Self-attention Does Not Need $O(n^2)$ Memory.

Modality Agnostic Efficient Long Range Encoder Self-attention Does Not Need $O(n^2)$ Memory

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.415528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.415528Z digest=sha256:4733eb8ef8894e4cff4e22c934c45b854ef0c501a9561996b851ebdcffea08f0

Observation ea68c8ca-08bd-443e-9ac4-bcea0df233a0 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

Modality Agnostic Efficient Long Range Encoder Learning Transferable Visual Models From Natural Language Supervision

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.419888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.419888Z digest=sha256:22474d6006338a04874dd9d7c107a713df77414205630d03b9226171e4d1927f

Observation cf490f3b-ba2d-4c84-bc08-b07d41e60581 · outbound

This paper cites Fast Transformer Decoding: One Write-Head is All You Need.

Modality Agnostic Efficient Long Range Encoder Fast Transformer Decoding: One Write-Head is All You Need

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.424102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.424102Z digest=sha256:5eeabf105ea78b9ead01e761965e8da3e2756a5ee320504456a941d2a9f37cdc

Observation 8e568b0c-f1df-468c-a9e3-1b02f0a5ae8c · outbound

This paper cites Dropout: A simple way to prevent neural networks from overfitting.

Modality Agnostic Efficient Long Range Encoder Dropout: A simple way to prevent neural networks from overfitting

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.015343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.428525Z digest=sha256:c4898aeec0f55dea577732cb875e5c7fcf3de85c45b71c623744f60621a00421

Observation d6b1d789-7f89-4778-bc55-3f9a920933e9 · outbound

This paper cites Scaling Granite Code Models to 128K Context.

Modality Agnostic Efficient Long Range Encoder Scaling Granite Code Models to 128K Context

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.433055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.433055Z digest=sha256:1451ddd2c5aabc0154095c80cf9d015306f7c1ef912786f5a2785df63567784b

Observation 8b34aa8f-5245-4abd-bf37-e9ac6d438833 · outbound

This paper cites Rethinking the inception architecture for computer vision.

Modality Agnostic Efficient Long Range Encoder Rethinking the inception architecture for computer vision

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:26.002442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.438136Z digest=sha256:bc6ad3fce23bc813ccbaf350218bc27ccc930c427a4e6c657713778b246d30d4

Observation 27c25218-4a4c-4e58-b892-c55857594733 · outbound

This paper cites Sparse sinkhorn attention.

Modality Agnostic Efficient Long Range Encoder Sparse sinkhorn attention

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.988808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.442255Z digest=sha256:93ead0be5e9ec22eadc9214cfcdd4f82794a9971f63792bd6903ca2a25d6f374

Observation c568d9a2-a81d-4380-97b2-11a0af045455 · outbound

This paper cites Long range arena: A benchmark for efficient transformers.

Modality Agnostic Efficient Long Range Encoder Long range arena: A benchmark for efficient transformers

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.975646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.446268Z digest=sha256:6561235b1d41a03a52b942f21f177adb5d5a4a8abad58a06df8b62d4352d66b1

Observation 9b3e551d-34f2-4704-b71c-0c858959ce64 · outbound

This paper cites Atten- tion is all you need.

Modality Agnostic Efficient Long Range Encoder Atten- tion is all you need

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.962476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.450471Z digest=sha256:5e1e7c9cd85f4ac1650adecfeb5898c138f332353c21582e72358676a4ad79d7

Observation 27f9710a-aefa-4135-8eab-6c0349fa9bd2 · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

Modality Agnostic Efficient Long Range Encoder Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.454731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.454731Z digest=sha256:530420e0b8dd1d74377a028e0393d3b73c30fbd3051bc7a91e711703affb6006

Observation d3cefd6b-9bfa-48fd-bdbe-9f1891eb5021 · outbound

This paper cites Huang, Krzysztof Choromanski, Valerii Likhosherstov, and Adrian Weller.

Modality Agnostic Efficient Long Range Encoder Huang, Krzysztof Choromanski, Valerii Likhosherstov, and Adrian Weller

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.948665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.459274Z digest=sha256:3d5f56777a447414a8f0becf885299bf0a878fc462d0b22ec53f4429ca0f5ec4

Observation 205864dc-13d0-4511-9f0b-658740b3868d · outbound

This paper cites MeSHup: A Corpus for Full Text Biomedical Document Indexing.

Modality Agnostic Efficient Long Range Encoder MeSHup: A Corpus for Full Text Biomedical Document Indexing

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:59:25.605672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.464664Z digest=sha256:89e7156f6ca73103d34db60d28ffc3047c266d25f709195a46d03b505753f628

Observation aed7673b-bfe3-4ea9-8b30-3d74e96e3ed5 · outbound

This paper cites Pytorch image models.

Modality Agnostic Efficient Long Range Encoder Pytorch image models

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.934518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.469293Z digest=sha256:9d2bf3f000ced3090bc2285415266b57f1a0ca55b0657ae071deef4d3ae2385a

Observation e5fe8155-0781-420d-a003-e7eaaacebf1c · outbound

This paper cites Ef- fective long-context scaling of foundation models.

Modality Agnostic Efficient Long Range Encoder Ef- fective long-context scaling of foundation models

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.920457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.473274Z digest=sha256:e9aae71ab37217095867039673c2fd2c0ff0b3b3d619dec708dadf44b932ce97

Observation 1da600db-4868-46d7-ac0d-f7ffa5f80bf4 · outbound

This paper cites A-vit: Adaptive tokens for efficient vision transformer.

Modality Agnostic Efficient Long Range Encoder A-vit: Adaptive tokens for efficient vision transformer

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.906773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.477372Z digest=sha256:c18519ef80df17b5b4f04a3cb0fcfec3b3cb7e6aff9ec877bb3c30eacd6452c7

Observation 65337461-5d63-4e10-b756-0c38205a0dbe · outbound

This paper cites MEGABYTE: Predicting Million-byte Sequences with Multiscale Transformers.

Modality Agnostic Efficient Long Range Encoder MEGABYTE: Predicting Million-byte Sequences with Multiscale Transformers

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.481718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.481718Z digest=sha256:60c969c83c5821012d06b54ee3610fb0205a35f9c1db6efbae9e3cb53437222a

Observation b1a65efb-5ebd-44f3-bf4f-7a6b80d8256e · outbound

This paper cites Metaformer is actually what you need for vision.

Modality Agnostic Efficient Long Range Encoder Metaformer is actually what you need for vision

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.892725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.486110Z digest=sha256:5eddba986aba97a1c978d8fe445ff6a66ef9d594a7d26be497ef89d0f146278a

Observation c4b9783c-bed3-400d-8f7c-dd4510e5d740 · outbound

This paper cites Cutmix: Regularization strategy to train strong classifiers with localizable features.

Modality Agnostic Efficient Long Range Encoder Cutmix: Regularization strategy to train strong classifiers with localizable features

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.877463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.490295Z digest=sha256:45be0a10835575b14121f4d54145935b6a67c19671826aaacc5cd0d1583534ff

Observation 1768df0d-4863-4051-8548-0f4155c2aea5 · outbound

This paper cites Big bird: Trans- formers for longer sequences.

Modality Agnostic Efficient Long Range Encoder Big bird: Trans- formers for longer sequences

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.863098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.494565Z digest=sha256:6e9ba672460ad1d5609b794a9d42903f5d6521961e42518a99ff341e2661f8ce

Observation c6651615-4d64-4f29-b1cc-0edb604dc96f · outbound

This paper cites mixup: Beyond empirical risk minimization.

Modality Agnostic Efficient Long Range Encoder mixup: Beyond empirical risk minimization

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.849604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.498610Z digest=sha256:c6e82180488b55adfa1754c12c013f83381b67ac4f6724855ac8c6438bf47c49

Observation 39cd3b71-1fcd-483c-8f13-9eaaa2c1e444 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

Modality Agnostic Efficient Long Range Encoder OPT: Open Pre-trained Transformer Language Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T17:59:25.502868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:59:25.502868Z digest=sha256:25120599aa311ca8cc0509d3423f84d2b9b994a0e4b014ef0fa12001ff93a23a

Observation 9db88336-7389-4943-966a-6d4c20ab901c · outbound

This paper cites Adavit: Adaptive vision transformers for efficient image recognition.

Modality Agnostic Efficient Long Range Encoder Adavit: Adaptive vision transformers for efficient image recognition

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.835846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.507154Z digest=sha256:27e36868e7815191fdbc36b9dc434e82d2e6d79809cb5f115772d1e1f876aec0

Observation 27b8126e-4ca5-4088-86e3-ad1eaa430b5f · outbound

This paper cites Long-short transformer: Efficient transformers for language and vision.

Modality Agnostic Efficient Long Range Encoder Long-short transformer: Efficient transformers for language and vision

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:59:25.822039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.511224Z digest=sha256:a615a1b7d37069541acbfb48df66830fe77b0c3f208c8c4f8b6f73c0e0f5dc1d

Observation feabdaa6-8dd1-42ba-857f-21e471ac9c98 · outbound

This paper cites Multiscale Audio Spectrogram Transformer for Efficient Audio Classification.

Modality Agnostic Efficient Long Range Encoder Multiscale Audio Spectrogram Transformer for Efficient Audio Classification

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-08-15T17:59:25.556777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T17:59:25.515340Z digest=sha256:15c3e901af782fd09ac11b123d17eaaf166a35891a05b21a86f2e29f3156e64e

Pith citing papers

No inbound Pith citation observations are available.