Pith. sign in

Paper Citation Record · LEDGER

Compressive Transformers for Long-Range Sequence Modelling

As of 6 August 2026, this Paper Citation Record lists 100 of 125 outbound references and 66 inbound Pith citation observations for arXiv:1911.05507.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1911.05507 v1

Coverage vector

measured 100 of 125 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-18T10:46:16.373197Z

measured 166 of 166 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 66 of 66 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:31:43.616971Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

100 of 125 outbound references displayed

  • verified exact4
  • verified fuzzy52
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch34

External citation measurements

49
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 14f1b0de-3282-41ad-b631-bb2f06cf9968 · outbound

This paper cites Al-Rfou, D.

Compressive Transformers for Long-Range Sequence Modelling Al-Rfou, D

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.756021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:99669f7d69efb9027678274944598f620f0fbbfb962b7a068f595ed73d7fac67

Observation 61dc761f-37a4-4016-929a-0908e78dfdaf · outbound

This paper cites an unresolved cited work.

Compressive Transformers for Long-Range Sequence Modelling Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-05-18T10:46:16.850932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:c434417ee2ad7cde77af6c0a16a51da337df9e86ce755bb55de8a5b17e7e39f8

Observation 2c55da32-9924-499b-8d5f-e34def6f7364 · outbound

This paper cites DeepMind Lab.

Compressive Transformers for Long-Range Sequence Modelling DeepMind Lab

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-18T10:46:16.698044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:b68206f46957bb545636a284cac550d00da6e9e41515a4730f58d638e710893d

Observation dde1abfd-f111-4915-9e3e-5b627c3860aa · outbound

This paper cites an unresolved cited work.

Compressive Transformers for Long-Range Sequence Modelling Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-18T10:46:16.854400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:c5e2982ae15edea2dd29eaca263941b1aef5c6f72da82c3d42f5a4210e1c0bce

Observation dc4127aa-e3ec-43a9-8782-d9ddbeab7655 · outbound

This paper cites Espeholt, H.

Compressive Transformers for Long-Range Sequence Modelling Espeholt, H

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.857718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:feb9bb18ba10f7ea2df2d7fe90be7a8fbda4c8408022b7764c70f0154e0b1b8f

Observation 3b71d78a-01bb-4a7d-a398-9119bfe12a5c · outbound

This paper cites Graves, G.

Compressive Transformers for Long-Range Sequence Modelling Graves, G

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.860846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:15c9416dff3ce66cc4c4a70a992b7be90d5f437a9b748c2007d19dfd3a8c2b05

Observation 62fc6e41-5c40-40a7-884c-3eb7f609dcb0 · outbound

This paper cites Hochreiter and J.

Compressive Transformers for Long-Range Sequence Modelling Hochreiter and J

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.864211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:f125c2c7d22b8a15017d1afa9637e9a14f1e7237fd4b2f08654a9e3100e37ac9

Observation ebfb2cbe-8a76-48f0-b098-48e22679645d · outbound

This paper cites an unresolved cited work.

Compressive Transformers for Long-Range Sequence Modelling Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-05-18T10:46:16.867165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:735632c0d1a47d80b55c96050bc48bc11c63ef0ef0cf8a9803767592b9ffc9d2

Observation 6cccd004-adcf-48e9-a663-9d2cd9b7691b · outbound

This paper cites Ko c isk \`y , J.

Compressive Transformers for Long-Range Sequence Modelling Ko c isk \`y , J

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.870112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:cd85abab8a9916c1b74cd0dd5d88ce848f1ea8a692657d84a2353d652221cd51

Observation 958832a4-d899-4a1a-a0e9-cdf1c8081547 · outbound

This paper cites Dynamic Evaluation of Transformer Language Models.

Compressive Transformers for Long-Range Sequence Modelling Dynamic Evaluation of Transformer Language Models

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-18T10:46:16.455743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:2b41f93012d8de7d770974effc98f4b86a46a00779fe8382ee269f43c65d22e8

Observation 0546c390-dc1d-48f6-b280-5f57cc8ea601 · outbound

This paper cites Mikolov, M.

Compressive Transformers for Long-Range Sequence Modelling Mikolov, M

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.872728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:04741e4d0c0d848b6582dfea337dede2426a82103a2eb54cb08842e22c131aa5

Observation 19850504-3b31-4c77-a03b-5e20c445462a · outbound

This paper cites an unresolved cited work.

Compressive Transformers for Long-Range Sequence Modelling Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-05-18T10:46:16.875648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:814b31be64fea63d0191c4144a0180dd85b3c670a35e3e48e1fba6f44fa372cf

Observation 951cbd9c-6036-47de-aa04-0de9f1de7a24 · outbound

This paper cites Paperno, G.

Compressive Transformers for Long-Range Sequence Modelling Paperno, G

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.878804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:e70bed9b68406250c8f88451bc53cdf71c627f64137654a5c5d78389cf7b1c23

Observation 90e48575-dfbf-4d52-8389-f7ff9aaa0827 · outbound

This paper cites an unresolved cited work.

Compressive Transformers for Long-Range Sequence Modelling Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-05-18T10:46:16.882023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:953c1dec214f1806868ec45fb14b73a5f282464e55e8b8bed03db7868db61bac

Observation 73a78e3d-4531-4e89-a06c-ae491972d450 · outbound

This paper cites an unresolved cited work.

Compressive Transformers for Long-Range Sequence Modelling Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-05-18T10:46:16.885212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:d520510f0cd15ff401c2386075327abf4f5f25bffd1b83e763f8d08818c4d9ff

Observation 0153efbc-eb91-4c02-8565-773a1ab5f2bc · outbound

This paper cites an unresolved cited work.

Compressive Transformers for Long-Range Sequence Modelling Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-05-18T10:46:16.888247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:abb7519a7acd7533e48ac92b6bb8ec71b9f605c175dd6a64966d5206af27439e

Observation 4d12252e-d82a-47a5-bf7d-375191e30f45 · outbound

This paper cites Santoro, R.

Compressive Transformers for Long-Range Sequence Modelling Santoro, R

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.891109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:1fe0c635bb75e07fa975200fcb0245add67e6a2f4e0d4806f14e4e017bcb92b7

Observation c699f14d-616e-4e77-8068-d363409da2e4 · outbound

This paper cites Shoeybi, M.

Compressive Transformers for Long-Range Sequence Modelling Shoeybi, M

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.894158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:65126612b24efb09aa02471b9f239d56b1626f705ec3f3424b6fdd91e8589dbb

Observation 85b01846-fd61-4f91-8ad2-642b459cdfc2 · outbound

This paper cites Smith, P.

Compressive Transformers for Long-Range Sequence Modelling Smith, P

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.897157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:a831d62a95ef4aff361503372227534590fa0c4f9f3cf8bf94f3bda362ca997d

Observation e7ec0fe2-12cc-42e3-9a92-ebc74b6f3427 · outbound

This paper cites Vaswani, N.

Compressive Transformers for Long-Range Sequence Modelling Vaswani, N

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.900541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:c59672463125f6c5f0da9f09cea00a744f3da54eb34c4684885293c97cfa130f

Observation 5d668fb3-e5e1-4baa-8a2b-ecdc36309eac · outbound

This paper cites an unresolved cited work.

Compressive Transformers for Long-Range Sequence Modelling Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-05-18T10:46:16.904485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:491830f8607102d032f46cc3e94b1fd85af4ed44baed2fb02ed7a58d09bc3b52

Observation bc74beb4-54d9-4d62-b5bd-41b13cd534a1 · outbound

This paper cites an unresolved cited work.

Compressive Transformers for Long-Range Sequence Modelling Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-05-18T10:46:16.907596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:adb45038ed15abfd7daa32e7a64ede77ca96c09e03f40e1e86a33ac7c8cf9da7

Observation e8b67db5-e87a-4274-9a2e-3e748dd8b033 · outbound

This paper cites an unresolved cited work.

Compressive Transformers for Long-Range Sequence Modelling Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-05-18T10:46:16.910545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:7312a0f46113c28328dd800c7f9593fc72c4e5ef38b976c29d23930644a782af

Observation 38f02672-c05f-436c-aa45-61efd8ea56c6 · outbound

This paper cites nature , volume=.

Compressive Transformers for Long-Range Sequence Modelling nature , volume=

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.914060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:7b5bfcde5adedf88015307af50d62233eaad72b14b5f86e8f095370b8d1b8c7e

Observation eb9b5a40-3308-4a47-969d-588e4018b647 · outbound

This paper cites Neural Turing Machines.

Compressive Transformers for Long-Range Sequence Modelling Neural Turing Machines

Reference 51

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.665814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:6d51c0a45552728846b1d5133342e6bc61d09ab8b7566fea791c62b983ca1ab4

Observation 8529a0f0-371a-4bc3-9963-774ffa63b442 · outbound

This paper cites Nature , volume=.

Compressive Transformers for Long-Range Sequence Modelling Nature , volume=

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.916851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:de8026d8a477bb7a5c932578847980be88c13b229039b441f105b9e7ca5ba718

Observation 9a3882bc-a404-4cd0-9bfc-01bfee36289c · outbound

This paper cites Advances in neural information processing systems , pages=.

Compressive Transformers for Long-Range Sequence Modelling Advances in neural information processing systems , pages=

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.920329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:b1c27b0afc8a8ecc1d88acf1a4bd538fde2328f536d5758eb12621810183e7da

Observation 25dc9822-30c2-4efe-b618-29c694940b93 · outbound

This paper cites Neural Machine Translation by Jointly Learning to Align and Translate.

Compressive Transformers for Long-Range Sequence Modelling Neural Machine Translation by Jointly Learning to Align and Translate

Reference 54

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.593979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:97a86270a639d839b5f42313e81f8b46c98df380ef643a29bbc568dd5ade4808

Observation 1a85df91-50f4-42ba-aa1f-1cab198bf8b1 · outbound

This paper cites Advances in neural information processing systems , pages=.

Compressive Transformers for Long-Range Sequence Modelling Advances in neural information processing systems , pages=

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.923246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:17eb092ba825c959dbc62c15d907564acb56ab7183b36f69d136ee496e522415

Observation 882d5cb8-e0e8-41ac-a591-41a3bf2b6104 · outbound

This paper cites Acoustics, speech and signal processing (icassp), 2013 ieee international conference on , pages=.

Compressive Transformers for Long-Range Sequence Modelling Acoustics, speech and signal processing (icassp), 2013 ieee international conference on , pages=

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.926426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:2e94c44c15326c8d5faaadeb635f844a70440c8d9fad968c6da81d58f3de1856

Observation 2375f3f9-6b14-4706-8226-6ff213e8cfae · outbound

This paper cites International Conference on Machine Learning , pages=.

Compressive Transformers for Long-Range Sequence Modelling International Conference on Machine Learning , pages=

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.929631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:1b7bca27c9fbc8bcc8ab3e0bc9d13faa7ee6ab15fde7702a585d17483dda54ca

Observation c65f45b4-f040-42fe-81a1-d0a2b31cb1d0 · outbound

This paper cites Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation.

Compressive Transformers for Long-Range Sequence Modelling Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation

Reference 58

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.531810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:9673441223ed00d82ec0d9b498a23db643b7b841d7b7f708fb0f363a9cd11ed6

Observation 4761d66e-d0e8-4345-851b-5444360e14f8 · outbound

This paper cites Neural computation , volume=.

Compressive Transformers for Long-Range Sequence Modelling Neural computation , volume=

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.932885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:2bdbacefcdd824b89702876496fd66481a4e5ed8ffa9ad4a208d3b1481403e44

Observation 2a604502-f44b-458c-aaf0-da1d1bbdf719 · outbound

This paper cites International Journal of Uncertainty, Fuzziness and Knowledge-Based Systems , volume=.

Compressive Transformers for Long-Range Sequence Modelling International Journal of Uncertainty, Fuzziness and Knowledge-Based Systems , volume=

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.937812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:47b8382e5e219f7744eedf8fd3c92d6b8cd6040c493473940287e17d91588da5

Observation b124e65c-d49d-4fe4-8c73-1d1c99137f87 · outbound

This paper cites Memory-based control with recurrent neural networks.

Compressive Transformers for Long-Range Sequence Modelling Memory-based control with recurrent neural networks

Reference 61

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.683229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:d23c9bf574fefbcb6b4571d13198341fbc23a5d8d7a22f807fbce241752079bf

Observation 80e2d20a-61ef-4a58-988a-8f6bf0141a89 · outbound

This paper cites One-shot Learning with Memory-Augmented Neural Networks.

Compressive Transformers for Long-Range Sequence Modelling One-shot Learning with Memory-Augmented Neural Networks

Reference 62

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.688673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:3094d59eef0eb4c820531796ebefa2f77857ef2044741948805342ff2cdd8557

Observation 89268fea-ca11-4c55-b745-a6b42d018f70 · outbound

This paper cites Advances in Neural Information Processing Systems , pages=.

Compressive Transformers for Long-Range Sequence Modelling Advances in Neural Information Processing Systems , pages=

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.940948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:85a409309eac05d0451e3863894d49bcc8dcb44b34711f119631e0383b4673c7

Observation ccf862c8-e024-49c4-b982-afb45c4c787a · outbound

This paper cites Advances in Neural Information Processing Systems , pages=.

Compressive Transformers for Long-Range Sequence Modelling Advances in Neural Information Processing Systems , pages=

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.947277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:eb7883edd05b92682834b154efc468a9fa1f645aa3bd8fc93b7d4f2c0bf4201b

Observation 964b98e2-22f5-4525-816a-78a01745fd9a · outbound

This paper cites Learning to Create and Reuse Words in Open-Vocabulary Neural Language Modeling.

Compressive Transformers for Long-Range Sequence Modelling Learning to Create and Reuse Words in Open-Vocabulary Neural Language Modeling

Reference 65

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.467292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:7306460df70fa98c00cfd170c903fe9ae1f23e2af15e7759c3ccc0496f30495d

Observation 13e2763d-33fa-4947-9525-c0a83c604698 · outbound

This paper cites Pointer Sentinel Mixture Models.

Compressive Transformers for Long-Range Sequence Modelling Pointer Sentinel Mixture Models

Reference 66

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.487194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:f63ae09b54a526242814332183df11006f09412ecb242fed00489119a8fd5b1f

Observation 3325e77b-d8a4-401e-96d7-0b27622976ad · outbound

This paper cites Improving Neural Language Models with a Continuous Cache.

Compressive Transformers for Long-Range Sequence Modelling Improving Neural Language Models with a Continuous Cache

Reference 67

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.517630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:8b6f69dbf34cd1aa697a969ab3d454476d98d68c8ecac1593e73e2720a009399

Observation fc0ccd2e-ed4c-4a3f-ad45-846c1b4c8941 · outbound

This paper cites Advances in Neural Information Processing Systems , pages=.

Compressive Transformers for Long-Range Sequence Modelling Advances in Neural Information Processing Systems , pages=

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.950693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:7f8403e7a0e4fa76005b83dcac851bd0c2f700ff1471f8d795b4d0c0350eed4a

Observation 41e4383c-375f-4908-b64b-961f2046ad67 · outbound

This paper cites Efficient softmax approximation for GPUs.

Compressive Transformers for Long-Range Sequence Modelling Efficient softmax approximation for GPUs

Reference 69

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.582529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:d39f016de76bef731f8c18ac82781c11a80d95744929b8bb6e7409df415981ec

Observation f85e9bc2-b424-4b3f-964b-08e577dcfcbb · outbound

This paper cites Acoustics, Speech, and Signal Processing.

Compressive Transformers for Long-Range Sequence Modelling Acoustics, Speech, and Signal Processing

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.955705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:a5ec3aa48c2e8ea309251bac1e9732adeea3ff1e1e9ff1b0a0dc7608f7b7a032

Observation 12c7aef1-be53-49c6-8305-aa917585e1ae · outbound

This paper cites Layer Normalization.

Compressive Transformers for Long-Range Sequence Modelling Layer Normalization

Reference 71

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.616698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:ee4aa5e29baf4117ec08c0aef705a7bea781cb600d0710e9eabda846d03d0ca7

Observation 2e206eb0-4850-4990-baa5-f6cb50edb57a · outbound

This paper cites Language Modeling with Gated Convolutional Networks.

Compressive Transformers for Long-Range Sequence Modelling Language Modeling with Gated Convolutional Networks

Reference 72

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.630593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:2922a362c01429a2586a7006d512c6f8fc9a12ca3f56f1544602bb133e501ead

Observation 233a4e9c-f36c-4ad1-ab62-cd865f88b96a · outbound

This paper cites Search Engine Guided Non-Parametric Neural Machine Translation.

Compressive Transformers for Long-Range Sequence Modelling Search Engine Guided Non-Parametric Neural Machine Translation

Reference 73

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.659674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:311cb0fd9ba2337fc13be92045f6725d0b929a572a63cef5f44f3c87389220d9

Observation 72e591f3-4bff-4034-ad5b-1a7c8715f4b8 · outbound

This paper cites Advances in Neural Information Processing Systems , pages=.

Compressive Transformers for Long-Range Sequence Modelling Advances in Neural Information Processing Systems , pages=

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.959551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:046cbab460409c90f1b7caec6c021fce86f4decc343f1635d5d1a4b944f209b8

Observation 897f60bb-516b-4aea-99bb-3b2be5fb50d0 · outbound

This paper cites Memory Aware Synapses: Learning what (not) to forget.

Compressive Transformers for Long-Range Sequence Modelling Memory Aware Synapses: Learning what (not) to forget

Reference 75

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.677716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:aa4311693bbc62056a0e0eb100f06bf9f65787a11cbff39de48265f255898693

Observation 4d5453cf-c78a-4862-b479-cdd66c53d4cc · outbound

This paper cites Acoustics, Speech, and Signal Processing.

Compressive Transformers for Long-Range Sequence Modelling Acoustics, Speech, and Signal Processing

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.963922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:e3b99bbe878cba97a679acf90f664a5c5758ad073c3da21e7d54f35b9fe6d998

Observation 3d9b5437-40b8-4a88-8ed3-553165042567 · outbound

This paper cites Thirteenth Annual Conference of the International Speech Communication Association , year=.

Compressive Transformers for Long-Range Sequence Modelling Thirteenth Annual Conference of the International Speech Communication Association , year=

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.966891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:948dc36e19c97b7d7c68548d0231dddfad96fbf57bba4776ddc2c16a6db1c9b8

Observation 17a09e04-bc2f-4f5b-ba7c-ba006de93d71 · outbound

This paper cites Proceedings of the 25th international conference on Machine learning , pages=.

Compressive Transformers for Long-Range Sequence Modelling Proceedings of the 25th international conference on Machine learning , pages=

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.971237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:2b93817d64ff46ea843217c30e1a12fa68b9e4dcbc1d16c8099375bae0a9a808

Observation 18da6790-1065-4e8b-bfaf-80818ee7abee · outbound

This paper cites A Convolutional Neural Network for Modelling Sentences.

Compressive Transformers for Long-Range Sequence Modelling A Convolutional Neural Network for Modelling Sentences

Reference 79

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.721929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:1956cad7809239da0c140b250696d80b5b1e09c83779d8dbf9d8687b6bc981ac

Observation ffcc4068-c12c-45df-ab8d-ed19f5deb328 · outbound

This paper cites Exploring the Limits of Language Modeling.

Compressive Transformers for Long-Range Sequence Modelling Exploring the Limits of Language Modeling

Reference 80

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.731598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:3ba1633ca07c542df9781170b1b8c6eb450bf01d207635576ff6a5c4fb496f0d

Observation c26a15f6-7e28-443c-9857-a70027020544 · outbound

This paper cites author=.

Compressive Transformers for Long-Range Sequence Modelling author=

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.973971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:c92c02902c9d2c150860c60fee31be469f6f9aafa9eb598d1c7b11836f2f87b5

Observation ba88e12a-f764-45a9-8d6f-cee15e27a9f8 · outbound

This paper cites Trends in cognitive sciences , volume=.

Compressive Transformers for Long-Range Sequence Modelling Trends in cognitive sciences , volume=

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.977580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:7e7fdfc582e5da9b3bc491573c3db8d4fbde3debd6c9bc716104f55a01bc1a75

Observation d57c507f-a40a-40b3-a976-3786bc2981da · outbound

This paper cites Pointing the Unknown Words.

Compressive Transformers for Long-Range Sequence Modelling Pointing the Unknown Words

Reference 83

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.474309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:31cea58fa31a13e2f2382ff32a9d0af1ed5cb878d2b52f6a1e6959530355437a

Observation 7be95c50-8231-475c-a504-7ccf20e81130 · outbound

This paper cites 1949 , publisher=.

Compressive Transformers for Long-Range Sequence Modelling 1949 , publisher=

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.980481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:6e22adcdf6405e0b277c8d6b9b112b28deec9d8fe1ec7a51bc2970605b333f6c

Observation cc90ed68-d6da-4f10-ad59-49636ee78f88 · outbound

This paper cites NY Houghton-Mifflin , year=.

Compressive Transformers for Long-Range Sequence Modelling NY Houghton-Mifflin , year=

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.983529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:b8ded9194eb14e1d0784e80697ca2563c79a4224d397f0fe24835fbbf8fad662

Observation 234ba166-f8c9-41a1-8533-f99cabe68bdb · outbound

This paper cites 2018 , url=.

Compressive Transformers for Long-Range Sequence Modelling 2018 , url=

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.987171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:d6ba8161a7eefc91674d080e75e57e66dc5d686b488ecfce5eb847be18ac9ba3

Observation b3ed0d0b-ab06-47a4-ae65-cf3945d1bd82 · outbound

This paper cites International Conference on Artificial Neural Networks , pages=.

Compressive Transformers for Long-Range Sequence Modelling International Conference on Artificial Neural Networks , pages=

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.990242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:4e9bfdb36d211efc6fc367b05d214427f727b0626d1131fd9e0649254e4995e9

Observation 22889612-60a6-4d08-b0f3-90e30b172d67 · outbound

This paper cites Advances in Neural Information Processing Systems , pages=.

Compressive Transformers for Long-Range Sequence Modelling Advances in Neural Information Processing Systems , pages=

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.993103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:648eb69c0a306b29a81830172643d66f9ca7c9b94d0394c37be7c2cae3bc4395

Observation 47506610-fd3d-4963-86c6-2751897bc113 · outbound

This paper cites Attentive Recurrent Comparators.

Compressive Transformers for Long-Range Sequence Modelling Attentive Recurrent Comparators

Reference 89

Resolution
verified exact
local_arxiv, observed 2026-05-18T10:46:16.609435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:703720d4afd1564336e1f6d33fbe92976567f05a59e7f05b5c569be716ef8fa9

Observation a01ed393-41bb-43c0-9a20-0a0995a2c9ed · outbound

This paper cites Proceedings of the National Academy of Sciences , volume=.

Compressive Transformers for Long-Range Sequence Modelling Proceedings of the National Academy of Sciences , volume=

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.996576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:7b19a00540f45835b9c87c51f05d1ea9afd0649f60e99a268bcb94636b1996e2

Observation d9e38191-e93b-463b-a72d-500dbb6f7d9c · outbound

This paper cites International Conference on Learning Representations , year=.

Compressive Transformers for Long-Range Sequence Modelling International Conference on Learning Representations , year=

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.999707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:d8197f40530b856c6d9e1a82bd2b6250a3d3b853013556d9d184710a8f676286

Observation 4c31bf8f-78fa-4d51-98ce-4f21e360447d · outbound

This paper cites Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks.

Compressive Transformers for Long-Range Sequence Modelling Model-Agnostic Meta-Learning for Fast Adaptation of Deep Networks

Reference 92

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.635514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:16ccb94eb3cfdb32ab4f270d11975aaa87916d0f19b5b1ef1539052aa9d398f9

Observation 85492747-8742-4698-9d48-ec5d57381028 · outbound

This paper cites Advances in neural information processing systems , pages=.

Compressive Transformers for Long-Range Sequence Modelling Advances in neural information processing systems , pages=

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:17.003392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:86c9e373113b682a999256359b8219fba0de131e5e973843fbf53d620de13bbf

Observation f42e9e0d-a4d2-4530-a268-b01390c4d07f · outbound

This paper cites Proceedings of the Thirteenth International Conference on Artificial Intelligence and Statistics , pages=.

Compressive Transformers for Long-Range Sequence Modelling Proceedings of the Thirteenth International Conference on Artificial Intelligence and Statistics , pages=

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.735605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:c19b84b6cdca132976507d5ad1f159f7574778ea3dcbf1925ca7ca6a79467903

Observation 0133baa6-6d43-4cdf-9782-05bca6773a71 · outbound

This paper cites Science , volume=.

Compressive Transformers for Long-Range Sequence Modelling Science , volume=

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.738966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:469bfc7a8501c3e048ae781082bae63cefe5507d87d2a6f27490b91b5664b3f1

Observation 4fef4839-c74d-4b43-8a10-61687642c36c · outbound

This paper cites COURSERA: Neural networks for machine learning , volume=.

Compressive Transformers for Long-Range Sequence Modelling COURSERA: Neural networks for machine learning , volume=

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.742377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:b2394cf964858ea1bdbc920079e23a0aed514bd99449cffe07efc5be6b607c34

Observation 9d08faa1-07c5-4c98-8b9f-10874c0a511d · outbound

This paper cites dvd , author=.

Compressive Transformers for Long-Range Sequence Modelling dvd , author=

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.745694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:17913cd8858f5b964088a942dc10e1f337834b3cc6e1524c0713f1e86496db17

Observation bf7b805a-1c64-4fbf-bb1a-9cf86a856411 · outbound

This paper cites On the State of the Art of Evaluation in Neural Language Models.

Compressive Transformers for Long-Range Sequence Modelling On the State of the Art of Evaluation in Neural Language Models

Reference 98

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.693638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:2961076daee00539325c6900da5265b3751eae813769a1110200c0b5950e9e9a

Observation e1289665-b79f-4dfb-b846-ecc41afb93dd · outbound

This paper cites 2017 , journal=.

Compressive Transformers for Long-Range Sequence Modelling 2017 , journal=

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.749725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:0d3f1356adaf9a2a368d03e1ce03fa2e3818f6d69ac2d0f1e9f688a4ac6e1d25

Observation 0c3c4725-eb75-4eba-8666-19750ea1484e · outbound

This paper cites Strategies for Training Large Vocabulary Neural Language Models.

Compressive Transformers for Long-Range Sequence Modelling Strategies for Training Large Vocabulary Neural Language Models

Reference 100

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.702975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:bed898e3d9b160fc8f283c10387e1f43dac45f6d3d42ac26b55d66f3042ae220

Observation 519e806b-9a87-4d32-9224-bec9e76eebef · outbound

This paper cites Deep Meta-Learning: Learning to Learn in the Concept Space.

Compressive Transformers for Long-Range Sequence Modelling Deep Meta-Learning: Learning to Learn in the Concept Space

Reference 101

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.707657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:db617ef669e25c27c6e1ebef4970fea52642b9196f0aa7855621b21d135c511f

Observation b2a9de35-17c8-4805-9611-676a9917d8c4 · outbound

This paper cites Breaking the Softmax Bottleneck: A High-Rank RNN Language Model.

Compressive Transformers for Long-Range Sequence Modelling Breaking the Softmax Bottleneck: A High-Rank RNN Language Model

Reference 102

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.717100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:831a7f8c8e030c37e7dd6f4cc53aa3fde010894dd2beb596ae0e7184c981efb9

Observation 8c008ac4-abde-46d9-9140-b5a4a9f3e4f0 · outbound

This paper cites Learning to learn , pages=.

Compressive Transformers for Long-Range Sequence Modelling Learning to learn , pages=

Reference 103

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.752962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:c6a19f13cbf1983f73e871c4bbdffeebc2bbbfe3e43590a602a46f6857f430e1

Observation d2916fc0-b51e-4a18-b84a-f67ff6c383e0 · outbound

This paper cites 1996a , school=.

Compressive Transformers for Long-Range Sequence Modelling 1996a , school=

Reference 104

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.847206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:ccddc98d68f580d9f0dfc8f9ff3973b590413ed51246ad798fab272f841ff145

Observation 149dee19-63c2-4f73-8c64-06ace70e7e9a · outbound

This paper cites Neural computation , volume=.

Compressive Transformers for Long-Range Sequence Modelling Neural computation , volume=

Reference 105

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.759195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:9fc95576b49265b7c57a2d698bc992c582ec66b0f81744778fda78cdd83b0645

Observation 62949997-c542-41ee-9b86-ee8a4f90a198 · outbound

This paper cites Advances in Neural Information Processing Systems , pages=.

Compressive Transformers for Long-Range Sequence Modelling Advances in Neural Information Processing Systems , pages=

Reference 106

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.762736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:1112e81a43ca1e89c52e417496970567dc2c753ba69271eafb5088f88cd490f6

Observation 8fd72273-9273-4191-92ff-9d82bf127091 · outbound

This paper cites Eleventh Annual Conference of the International Speech Communication Association , year=.

Compressive Transformers for Long-Range Sequence Modelling Eleventh Annual Conference of the International Speech Communication Association , year=

Reference 107

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.766652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:4abf07bf4c2861b7d266498c96588bbae60526c8c742837b32072d4fffe13f9a

Observation 914581d2-d3d4-4dcf-8cc4-601c20718b95 · outbound

This paper cites Fast Parametric Learning with Activation Memorization.

Compressive Transformers for Long-Range Sequence Modelling Fast Parametric Learning with Activation Memorization

Reference 108

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.480480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:c13426455539f1bc3bd0526c4a887d1b9210b4a83408be8401cdfa99b74eb76d

Observation f4a208bb-7e8c-4845-abf2-71b3836247f9 · outbound

This paper cites Advances in Neural Information Processing Systems , pages=.

Compressive Transformers for Long-Range Sequence Modelling Advances in Neural Information Processing Systems , pages=

Reference 109

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.770232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:b14ac495b048021df2cd8ebbaac6b827f701086df3a4a4f4f4b81c18645d10a8

Observation b1ea3cc4-a4a6-4b64-a11f-bd5bff1e80b5 · outbound

This paper cites Adaptive Input Representations for Neural Language Modeling.

Compressive Transformers for Long-Range Sequence Modelling Adaptive Input Representations for Neural Language Modeling

Reference 110

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.493965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:68df3f32b559e3ce4eb05697ecb8640a877e49f95a6c6761e65caac565471745

Observation 5215501a-4072-49da-99be-20fe95e04ced · outbound

This paper cites Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context.

Compressive Transformers for Long-Range Sequence Modelling Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context

Reference 111

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.499561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:14bcde58fafdd4956b901b063c600bf14063f48ddf8ae6df0a9f0fe23e51a96a

Observation 200a289e-cbe0-4cc0-96ee-74d22302b029 · outbound

This paper cites HyperNetworks.

Compressive Transformers for Long-Range Sequence Modelling HyperNetworks

Reference 112

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.504786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:35d20f5df6051de0f8b4258ee1a77ee74f47c762f03439c956faf9ea18f9c599

Observation bcd11c4f-fdb0-450a-8519-36b8c067fdfa · outbound

This paper cites Hierarchical Multiscale Recurrent Neural Networks.

Compressive Transformers for Long-Range Sequence Modelling Hierarchical Multiscale Recurrent Neural Networks

Reference 113

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.511171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:85ff997bbb27828060e0abd9150c3e56f32d75aa0b256b8450b4c914c216f877

Observation 03e0ad89-e809-4a3b-805c-2b959edc3252 · outbound

This paper cites Proceedings of the 34th International Conference on Machine Learning-Volume 70 , pages=.

Compressive Transformers for Long-Range Sequence Modelling Proceedings of the 34th International Conference on Machine Learning-Volume 70 , pages=

Reference 114

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.774201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:63743d488f38519d78deeb86668fa184bd6967e5f7d1cdc23cdf1b1a4efcb0d2

Observation 549acdf6-fd6e-4043-8279-d6f8fe166dd2 · outbound

This paper cites Multiplicative LSTM for sequence modelling.

Compressive Transformers for Long-Range Sequence Modelling Multiplicative LSTM for sequence modelling

Reference 115

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.524083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:095f75878ea391b0aed4189d9e3dc46267ddf6f8b9e65775e5c9ef51c1979cf8

Observation e904083a-3c20-4f5d-8aae-72181104069c · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Compressive Transformers for Long-Range Sequence Modelling Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 116

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.778718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:ed2206fd88cd43ad5b252334f1ed4fcf1d548ef2e17813522e583ea4206201db

Observation 50901ec4-d001-44db-bac7-c5874fc068fd · outbound

This paper cites Generating Sequences With Recurrent Neural Networks.

Compressive Transformers for Long-Range Sequence Modelling Generating Sequences With Recurrent Neural Networks

Reference 117

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.537878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:62f269c911a24f1d7a32e44eb98bc84f5f2164b51a6e79deff69d550a666b66d

Observation e6374931-4216-4b32-a97c-e6e38396bd45 · outbound

This paper cites Neural Machine Translation in Linear Time.

Compressive Transformers for Long-Range Sequence Modelling Neural Machine Translation in Linear Time

Reference 118

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.543617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:1c7914a1cfa349a2694122796310ae3d1ab5b7aef7815e6df168e869ce260977

Observation f9f4e484-35c2-4926-9170-77614067a27f · outbound

This paper cites Quasi-Recurrent Neural Networks.

Compressive Transformers for Long-Range Sequence Modelling Quasi-Recurrent Neural Networks

Reference 119

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.549503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:bc3bd1947a36589cbbe8875b4934a971382fe2a16f23951c9af066598117f503

Observation 9e64f01b-3e4a-42a1-ac60-f908c93aaf62 · outbound

This paper cites Generating Long Sequences with Sparse Transformers.

Compressive Transformers for Long-Range Sequence Modelling Generating Long Sequences with Sparse Transformers

Reference 120

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.556074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:ddd4e472a4243b84c3d4fba5df0168d77e0affddfb6191111aa9cdf2a8a15769

Observation f9bcd4de-875f-426d-bffa-9045c64bd324 · outbound

This paper cites Adaptive Attention Span in Transformers.

Compressive Transformers for Long-Range Sequence Modelling Adaptive Attention Span in Transformers

Reference 121

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:46:16.561868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:45235a94768c76522a0641ee40be8e5abf83d5af5df533a3b14999fa70ad62d4

Observation 5992963e-c0bd-463e-a49a-80b6e17977aa · outbound

This paper cites One Billion Word Benchmark for Measuring Progress in Statistical Language Modeling.

Compressive Transformers for Long-Range Sequence Modelling One Billion Word Benchmark for Measuring Progress in Statistical Language Modeling

Reference 122

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.570709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:0f11addde8e664003083e8354522b28b933f63ff06dd96a3f53f8e522aefa427

Observation 66167163-2d6b-4229-b232-7f097e54cf6a · outbound

This paper cites Trellis Networks for Sequence Modeling.

Compressive Transformers for Long-Range Sequence Modelling Trellis Networks for Sequence Modeling

Reference 123

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.576678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:a4f37f6e9d97888cdd1cfbb5f26eb8d86496649633850c40cbb2dd274c99d9d0

Observation b28a4541-61b4-457d-a8c7-e601d94543f2 · outbound

This paper cites 2016 , organization=.

Compressive Transformers for Long-Range Sequence Modelling 2016 , organization=

Reference 124

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.782636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:c9d24135a63a40b74bb55c70a20e67d82aa9dc0fd04a3c427ec688e06fccbf5c

Observation 4067da91-de35-4d1a-8933-856c062ce62b · outbound

This paper cites The Goldilocks Principle: Reading Children's Books with Explicit Memory Representations.

Compressive Transformers for Long-Range Sequence Modelling The Goldilocks Principle: Reading Children's Books with Explicit Memory Representations

Reference 125

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T10:46:16.587737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:36c2c896c61eb4b8cb8e7060639129200b5fcb3227fd15a8356e40cc7dea433e

Observation cd39ff37-907b-4a61-a80e-696c7feda8bb · outbound

This paper cites Proceedings of the IEEE international conference on computer vision , pages=.

Compressive Transformers for Long-Range Sequence Modelling Proceedings of the IEEE international conference on computer vision , pages=

Reference 126

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T10:46:16.786410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T10:46:16.373197Z digest=sha256:1f213a8222f1899f980977150f03913a77dd41112a8246497da05aa4dc643910

Pith citing papers

Observation 63809903-8635-4bd3-b81f-8a7560c4b954 · inbound

The Pile: An 800GB Dataset of Diverse Text for Language Modeling cites this paper.

The Pile: An 800GB Dataset of Diverse Text for Language Modeling Compressive Transformers for Long-Range Sequence Modelling

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T21:35:18.513342Z digest=sha256:02edb7a31e34e7d58443dac495d3217a622d2834f57db5fbd38c71f99efc4a1f

Observation 472d0823-1309-469b-b33b-ad06489daa07 · inbound

GPT-NeoX-20B: An Open-Source Autoregressive Language Model cites this paper.

GPT-NeoX-20B: An Open-Source Autoregressive Language Model Compressive Transformers for Long-Range Sequence Modelling

Reference 74

Resolution
verified exact
local_arxiv, observed 2026-05-24T12:34:28.437554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-24T12:33:37.701655Z digest=sha256:051f2f0c90db4fe3d4c9e400dea9602d5371bfb181dd549c21a72194feb6b9e0

Observation 3a48f80c-fad7-40f5-980d-001e1e13e499 · inbound

LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens cites this paper.

LongRoPE: Extending LLM Context Window Beyond 2 Million Tokens Compressive Transformers for Long-Range Sequence Modelling

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T08:31:04.774135Z digest=sha256:dfda6dc704c46b9f1aeb9447161dfd37657660b3da067efe5b5396c7dbb2331c

Observation 7fa00ae7-6ad3-488a-b6fa-c8386789d63d · inbound

Massive Activations in Large Language Models cites this paper.

Massive Activations in Large Language Models Compressive Transformers for Long-Range Sequence Modelling

Reference 145

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-16T07:02:53.740597Z digest=sha256:2bc994aecabe293cff8caaf4459d8f8198ac9698cc1c58170c438e83ab847c3a

Observation 2dd8d092-b481-41e2-869a-ff31dc0309a1 · inbound

Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention cites this paper.

Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention Compressive Transformers for Long-Range Sequence Modelling

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-21T18:17:00.268630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T18:17:00.157129Z digest=sha256:b9bc9e8863810f148866e1550f3420e69a1516f50d2b0fc8dd83f2cd40b7a646

Observation b8f20cf7-42a5-479a-8dac-bdd9c725e96f · inbound

Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference cites this paper.

Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference Compressive Transformers for Long-Range Sequence Modelling

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-15T14:12:20.809594Z digest=sha256:87cd489129463398517caacda819640eefffe67ae9442bd4c3376e607a62177d

Observation 2eca5e19-4bec-41b5-aa16-ec4782087a75 · inbound

Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach cites this paper.

Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach Compressive Transformers for Long-Range Sequence Modelling

Reference 125

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-12T15:39:40.845703Z digest=sha256:ddcb441e5712ae9aa8b0d9fe8e53a5c031bcea901557833a04d5406dc82a2a6c

Observation 46ce69cf-37df-4b9a-b60b-22475a9d62bc · inbound

Sparse Attention Remapping with Clustering for Efficient LLM Decoding on PIM cites this paper.

Sparse Attention Remapping with Clustering for Efficient LLM Decoding on PIM Compressive Transformers for Long-Range Sequence Modelling

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-22T17:01:48.565577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T16:58:11.105511Z digest=sha256:7b465bdc00456e1eb0256d8a6811645a7f10bb927779398e1d4d20b75f820501

Observation cdcaf6aa-a5b6-4586-b7a8-9ab090614695 · inbound

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems cites this paper.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Compressive Transformers for Long-Range Sequence Modelling

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:43.616971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:43.616971Z digest=sha256:67cf6da8e199d562d9a3b0a0ead651f92c0498b214e08d8b91cd615f897768c3

Observation 948eb287-38f8-4ba1-aeff-1bbcd9720208 · inbound

Memory-Augmented Transformers: A Systematic Review from Neuroscience Principles to Enhanced Model Architectures cites this paper.

Memory-Augmented Transformers: A Systematic Review from Neuroscience Principles to Enhanced Model Architectures Compressive Transformers for Long-Range Sequence Modelling

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-05T20:15:53.765855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:15:53.765855Z digest=sha256:f58ed86a42728a5a0cafa9adddfc91733ca693e18e3558de1644620b8e970fb9

Observation adb9c764-4364-469c-80c3-c8b0d9920fde · inbound

SemToken: Semantic-Aware Tokenization for Efficient Long-Context Language Modeling cites this paper.

SemToken: Semantic-Aware Tokenization for Efficient Long-Context Language Modeling Compressive Transformers for Long-Range Sequence Modelling

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T18:07:40.090807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:07:40.090807Z digest=sha256:7ca6bdd9b5c2f4900dfd04a0f001756f2de4666388f853ef2263d2473694059c

Observation 4f1579aa-a5dc-45f3-b2ab-3964f193006d · inbound

Quantum-Enhanced Optimization by Warm Starts cites this paper.

Quantum-Enhanced Optimization by Warm Starts Compressive Transformers for Long-Range Sequence Modelling

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T17:26:02.290468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:26:02.290468Z digest=sha256:bc8fa992781b99fc1ce87413c60cd3e8181c371d8e02b9b517377a22be42c66f

Observation 48d4fc6e-9e8f-40bd-acd3-ceb99b965205 · inbound

Vision encoders should be image size agnostic and task driven cites this paper.

Vision encoders should be image size agnostic and task driven Compressive Transformers for Long-Range Sequence Modelling

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T17:27:27.408925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:27:27.408925Z digest=sha256:2835dc5f86fe02aeec173afa5a9c449f7d804023723de30e4b87aaf9c80fe7c7

Observation 3a4fd09a-84f8-4d0f-89ec-3a2e57246d41 · inbound

Beyond Memorization: Extending Reasoning Depth with Recurrence, Memory and Test-Time Compute Scaling cites this paper.

Beyond Memorization: Extending Reasoning Depth with Recurrence, Memory and Test-Time Compute Scaling Compressive Transformers for Long-Range Sequence Modelling

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-05-18T20:51:50.848944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T20:49:26.966293Z digest=sha256:e82acc80306456a6b097932616afd21899d09d1411a2705b6f2e5eef0b5f8f35

Observation 3de31226-4896-40b5-8dd7-9044c00dd5ff · inbound

HoPE: Hyperbolic Rotary Positional Encoding for Stable Long-Range Dependency Modeling in Large Language Models cites this paper.

HoPE: Hyperbolic Rotary Positional Encoding for Stable Long-Range Dependency Modeling in Large Language Models Compressive Transformers for Long-Range Sequence Modelling

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T05:36:55.106978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:36:55.106978Z digest=sha256:f55ca75c865d5bf5f95f9afbe2571aadb4c97836bb3dacc953b5695889ea33e3

Observation 5d869398-4cd6-4fc4-a352-9d5bf6659e52 · inbound

Positional Encoding via Token-Aware Phase Attention cites this paper.

Positional Encoding via Token-Aware Phase Attention Compressive Transformers for Long-Range Sequence Modelling

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T16:21:36.681588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T16:19:48.318702Z digest=sha256:d935abad2fa5e838305c1ae5deeb96af628b38ae3dbb3a11fc4e455083835392

Observation c08777be-abd3-4ce6-bf49-45963a88d071 · inbound

ChunkLLM: A Lightweight Pluggable Framework for Accelerating LLMs Inference cites this paper.

ChunkLLM: A Lightweight Pluggable Framework for Accelerating LLMs Inference Compressive Transformers for Long-Range Sequence Modelling

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T14:43:03.696218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:43:03.696218Z digest=sha256:8b18704034297822350b34149829ca23476e8c1f37c5cdf6df1da7b2902110be

Observation e6935100-463b-4ad1-9f3f-3d4eae79cbcb · inbound

Hybrid Architectures for Language Models: Systematic Analysis and Design Insights cites this paper.

Hybrid Architectures for Language Models: Systematic Analysis and Design Insights Compressive Transformers for Long-Range Sequence Modelling

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T10:18:04.431436Z digest=sha256:60b6505764595f2290ee601f69ee7e99b9413a3647dd8c905784c7344a950dc9

Observation 3fafd378-c5e3-4f4c-8aea-981b1e5929ff · inbound

OBCache: Optimal Brain KV Cache Pruning for Efficient Long-Context LLM Inference cites this paper.

OBCache: Optimal Brain KV Cache Pruning for Efficient Long-Context LLM Inference Compressive Transformers for Long-Range Sequence Modelling

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T10:57:43.845157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:57:43.845157Z digest=sha256:fca7aaccbe9ae7bd8277dde961ef18615779764cf638b3afe972b0771037c886

Observation 9acf4671-7194-4cb0-a65e-e8227b7d8ab9 · inbound

MTraining: Distributed Dynamic Sparse Attention for Efficient Ultra-Long Context Training cites this paper.

MTraining: Distributed Dynamic Sparse Attention for Efficient Ultra-Long Context Training Compressive Transformers for Long-Range Sequence Modelling

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-21T19:44:19.647760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:44:04.833504Z digest=sha256:fc695040b081369cea0b8a6098fce478e6632940016ca806f7a5535054ddf533

Observation a47257d8-5a90-4870-aead-a7dd9d5bb1fd · inbound

ARC-Encoder: learning compressed text representations for large language models cites this paper.

ARC-Encoder: learning compressed text representations for large language models Compressive Transformers for Long-Range Sequence Modelling

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T08:28:51.749218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:28:51.749218Z digest=sha256:1e8bb6952e16f138e040ad76424b2ff6402031793622078f44b75baddcea5cc8

Observation acf31256-ba86-4ec8-b50d-f1bda63de08e · inbound

Beyond Context: Large Language Models' Failure to Grasp Users' Intent cites this paper.

Beyond Context: Large Language Models' Failure to Grasp Users' Intent Compressive Transformers for Long-Range Sequence Modelling

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:09:25.827452Z digest=sha256:dcc693029504448744c1becc7d1688bd53655b2503174f10a036f67e58a7374f

Observation 3e114493-9e38-4368-b471-f3ec41da3ceb · inbound

Prism: Spectral-Aware Block-Sparse Attention cites this paper.

Prism: Spectral-Aware Block-Sparse Attention Compressive Transformers for Long-Range Sequence Modelling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T03:22:54.245744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:22:54.245744Z digest=sha256:e2f58f95f02075bdc370c674ecda69852c34cb12ef9b663bf291607cd0c943c4

Observation 59f0eb41-f2e9-4efd-b797-d8ac2e434d37 · inbound

LPC-SM: Local Predictive Coding and Sparse Memory for Long-Context Language Modeling cites this paper.

LPC-SM: Local Predictive Coding and Sparse Memory for Long-Context Language Modeling Compressive Transformers for Long-Range Sequence Modelling

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T11:22:08.937342Z digest=sha256:a60e622ec1600fddf82ffc6c5f9c7ee82e472155f9d458c3c3b4c7aa20fd1958

Observation 7fd2fd39-ee49-4da2-bb98-43e7956f4c88 · inbound

Next-Scale Autoregressive Models for Text-to-Motion Generation cites this paper.

Next-Scale Autoregressive Models for Text-to-Motion Generation Compressive Transformers for Long-Range Sequence Modelling

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T17:27:34.972785Z digest=sha256:fe6cc798723670e46a466d3d6f0c584162350c76dc6b3946b450f43e96fb8cc8

Observation a1dedfaf-8d67-471f-b182-e03222374a82 · inbound

Next-Scale Autoregressive Models for Text-to-Motion Generation cites this paper.

Next-Scale Autoregressive Models for Text-to-Motion Generation Compressive Transformers for Long-Range Sequence Modelling

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-13T12:16:25.843631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:16:25.843631Z digest=sha256:5d962688ae9f9052d4e2735eef326bd4b00ecd77b5022a87b415df7c8ed7aa98

Observation 1906eeb9-07dc-4f4a-8170-6e268ef191a1 · inbound

StructKV: Preserving the Structural Skeleton for Scalable Long-Context Inference cites this paper.

StructKV: Preserving the Structural Skeleton for Scalable Long-Context Inference Compressive Transformers for Long-Range Sequence Modelling

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T18:37:10.526359Z digest=sha256:095d5d5fa3e6166bff11db9e99d01d7562e5bca29b6f1ea5cc0c30dfea481e3a

Observation c222a16f-f403-4377-8cb9-746b3e654eb7 · inbound

Training Transformers for KV Cache Compressibility cites this paper.

Training Transformers for KV Cache Compressibility Compressive Transformers for Long-Range Sequence Modelling

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T14:20:34.452801Z digest=sha256:e77f65089a3d0ec086e3a5b59e89a78ff045b91e8a1bd17a45cb8f117e3abdab

Observation 4c76dc23-6216-4dfa-b833-1a6b6a2f153f · inbound

Training Transformers for KV Cache Compressibility cites this paper.

Training Transformers for KV Cache Compressibility Compressive Transformers for Long-Range Sequence Modelling

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T06:01:40.766843Z digest=sha256:dd5b9833fac5c11da0e549220c4218cc5fa54812c261560ce06ad266444f0eff

Observation 4634c71a-bf3b-436d-a5cf-4bef74f422f2 · inbound

Adaptive Memory Decay for Log-Linear Attention cites this paper.

Adaptive Memory Decay for Log-Linear Attention Compressive Transformers for Long-Range Sequence Modelling

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T01:02:56.848785Z digest=sha256:37eaa911f95041cda55f2a9bb20e1ca9be5dba83088e5706e04014632ed8f340

Observation 0272af1d-c476-4a06-90de-0c161292e790 · inbound

Self-Consolidating Language Models: Continual Knowledge Incorporation from Context cites this paper.

Self-Consolidating Language Models: Continual Knowledge Incorporation from Context Compressive Transformers for Long-Range Sequence Modelling

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T01:47:15.491588Z digest=sha256:ee93ca8b29c4eb4f0da0fbda05cd154273176e3edf86294c9e0f63a462eca4ff

Observation 452525e6-2079-4b5c-956e-2c987be69c1c · inbound

Self-Consolidating Language Models: Continual Knowledge Incorporation from Context cites this paper.

Self-Consolidating Language Models: Continual Knowledge Incorporation from Context Compressive Transformers for Long-Range Sequence Modelling

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:53:43.047696Z digest=sha256:6401f900be6fa911fa1194460da17ee05da4beefb3b19ac791a1f3aefc852ebe

Observation d00a0907-a585-4fe3-a6dc-ebaad3b76fcb · inbound

A Unified Framework for Critical Scaling of Inverse Temperature in Self-Attention cites this paper.

A Unified Framework for Critical Scaling of Inverse Temperature in Self-Attention Compressive Transformers for Long-Range Sequence Modelling

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-14T19:56:02.462786Z digest=sha256:f71fcc34145160fb8d9d8aef199cfa3552ad77b8b910e1159edbaa9609d74605

Observation 01ca2394-6f6f-4412-ae37-0027b9767b96 · inbound

Phasor Memory Networks: Stable Backpropagation Through Time for Scalable Explicit Memory cites this paper.

Phasor Memory Networks: Stable Backpropagation Through Time for Scalable Explicit Memory Compressive Transformers for Long-Range Sequence Modelling

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T20:36:12.321162Z digest=sha256:06948df9cf096b609da9ef68652ffeb05d8758134c9baa97cab9768b5e05fcec

Observation 160af0ae-55de-4745-a0d8-725a7434f862 · inbound

OSDN: Improving Delta Rule with Provable Online Preconditioning in Linear Attention cites this paper.

OSDN: Improving Delta Rule with Provable Online Preconditioning in Linear Attention Compressive Transformers for Long-Range Sequence Modelling

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T19:08:18.768344Z digest=sha256:39fbf13291c227ca4bfffd9577cfb30345f7ba623187914f4979731d9ded31d8

Observation 4bb91e15-a8a5-4d08-b720-8bb0ccb348b7 · inbound

Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility cites this paper.

Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility Compressive Transformers for Long-Range Sequence Modelling

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:46:17.004740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-15T05:35:09.705532Z digest=sha256:0e320c882fa9fbaa6b252e739aa3c012c7d9b67d87ecd5f4c29881a1f537cdd8

Observation 9f22a9a7-4c73-450d-b969-c2053966c400 · inbound

Memory-Augmented Query Intent Understanding for Efficient Chat-based Image Retrieval cites this paper.

Memory-Augmented Query Intent Understanding for Efficient Chat-based Image Retrieval Compressive Transformers for Long-Range Sequence Modelling

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-20T13:33:19.119023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T13:31:45.891242Z digest=sha256:90926637d001a7e05301ffde80a8a84238cbfb36a493b32295c16f3af2dc4d05

Observation 968d7557-c88c-437e-8442-96e931cf5092 · inbound

SPHERICAL KV: Angle-Domain Attention and Rate-Distortion Retention for Efficient Long-Context Inference cites this paper.

SPHERICAL KV: Angle-Domain Attention and Rate-Distortion Retention for Efficient Long-Context Inference Compressive Transformers for Long-Range Sequence Modelling

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T20:33:43.292296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T20:30:33.208386Z digest=sha256:f37344d25048d8c59c41eddba61725ed3672cc234713a61e7af164daa848a1c3

Observation e40f18a9-e2ff-425f-a0e8-e4b29d91d6ee · inbound

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression cites this paper.

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression Compressive Transformers for Long-Range Sequence Modelling

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T05:11:06.148819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T05:10:53.133346Z digest=sha256:95b9cdeada5a073e78a0120e2e39ad528e52bdd97e9419ba407830c43e827ff3

Observation c7adbd75-c8d7-4abc-bfce-fb4d2991447a · inbound

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression cites this paper.

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression Compressive Transformers for Long-Range Sequence Modelling

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T17:34:57.862128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:28:08.750049Z digest=sha256:675f7efd8d1083af1c4db3c84fcdd6b4bb12a0634fbc8eeb258dcda8ba6fe4cc

Observation d71b0822-891c-4e53-8020-282e8eb4da50 · inbound

Tensor Cache: Eviction-conditioned Associative Memory for Transformers cites this paper.

Tensor Cache: Eviction-conditioned Associative Memory for Transformers Compressive Transformers for Long-Range Sequence Modelling

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T05:50:23.541466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-25T05:49:43.160574Z digest=sha256:a31c0578b9fbb779f9b22be7c771eccae7441309abdaa27c28c87573d7ce38ca

Observation eb9e9bd5-3a79-4c8d-9263-11c7c44f33b2 · inbound

Latent Recurrent Transformer: Architecture Exploration, Training Strategies, and Scaling Behavior cites this paper.

Latent Recurrent Transformer: Architecture Exploration, Training Strategies, and Scaling Behavior Compressive Transformers for Long-Range Sequence Modelling

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-06-29T19:43:54.972100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T19:36:13.559393Z digest=sha256:e06b5546379ebe4e2d1e75ff46bbaf9fe5ba7ad27ebded99280ccccc731b90b2

Observation ba73e450-642d-4a8d-a84a-a5e9f8d1aa74 · inbound

Lost in Sampling: Assessing Lexical Reachability in LLMs via the Word Coverage Score (WCS) cites this paper.

Lost in Sampling: Assessing Lexical Reachability in LLMs via the Word Coverage Score (WCS) Compressive Transformers for Long-Range Sequence Modelling

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T18:13:48.865948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T18:07:58.776272Z digest=sha256:101194ce26633dbc13f77182b75ed08f1368e2b2c5f74ee381e098a03d44d5f0

Observation e4dd4fd4-4db2-4eed-98c7-7feb03a7270b · inbound

Tensor Memory: Fixed-Size Recurrent State for Long-Horizon Transformers cites this paper.

Tensor Memory: Fixed-Size Recurrent State for Long-Horizon Transformers Compressive Transformers for Long-Range Sequence Modelling

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T18:13:48.720885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T18:08:42.533804Z digest=sha256:86b0d86cbe2b5116d7250ba0cfaeffb5426a4c5d133ebee246e11bb53d48e6cc

Observation 02102b0d-65be-4540-94bc-e708a9ca550a · inbound

Memory by Design: Probabilistic Sequence Layers cites this paper.

Memory by Design: Probabilistic Sequence Layers Compressive Transformers for Long-Range Sequence Modelling

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T20:26:12.837033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-28T21:07:31.407554Z digest=sha256:498840f4dd356a72630218cb739200fa8cd82a14de67a55283f664076bb7dfdb

Observation d3deb08a-3283-4f09-bab6-bfa64733516d · inbound

SHARP: Sleep-based Hierarchical Accelerated Replay for Long Range Non-Stationary Temporal Pattern Recognition cites this paper.

SHARP: Sleep-based Hierarchical Accelerated Replay for Long Range Non-Stationary Temporal Pattern Recognition Compressive Transformers for Long-Range Sequence Modelling

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-28T20:22:37.816854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T18:43:49.689283Z digest=sha256:682417f3d3eb8ac6fbabf4b37bb571df4d4634d2e8dab7018f2850c93fa28eeb

Observation bc30eaef-04f8-4010-a89b-8ebf85687cbf · inbound

SHARP: Sleep-based Hierarchical Accelerated Replay for Long Range Non-Stationary Temporal Pattern Recognition cites this paper.

SHARP: Sleep-based Hierarchical Accelerated Replay for Long Range Non-Stationary Temporal Pattern Recognition Compressive Transformers for Long-Range Sequence Modelling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T15:29:20.475050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T15:29:20.475050Z digest=sha256:6dd31e4071c946032e452b2d0b38f655c79019f77f49c883fec5d656ad05f51d

Observation f0656ac3-a417-40a1-8378-729ac1ad01c9 · inbound

SHARP: Sleep-based Hierarchical Accelerated Replay for Long Range Non-Stationary Temporal Pattern Recognition cites this paper.

SHARP: Sleep-based Hierarchical Accelerated Replay for Long Range Non-Stationary Temporal Pattern Recognition Compressive Transformers for Long-Range Sequence Modelling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-13T07:45:49.272396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T07:45:49.272396Z digest=sha256:e632bba2bbe052e5f5c8d2d890e38d0494ed16b28e8cd0f730162a08aea4511e

Observation 125aa08e-2c4c-40ba-ba37-434fb90e72e5 · inbound

Transformer-Enhanced Reinforcement Learning: Fundamentals and Applications in Communication Networks cites this paper.

Transformer-Enhanced Reinforcement Learning: Fundamentals and Applications in Communication Networks Compressive Transformers for Long-Range Sequence Modelling

Reference 82

Resolution
verified exact
local_arxiv, observed 2026-06-29T16:23:39.681605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T16:14:25.741686Z digest=sha256:fa9fab817d6a0df4008c23f8416ab200c230e413ddf87ea87bfa19b78043aecb

Observation 54cd3937-2aa4-4ccc-92e7-7b9ae87ec8b2 · inbound

TextEconomizer: Enhancing Lossy Text Compression with Denoising Transformers and Entropy Coding cites this paper.

TextEconomizer: Enhancing Lossy Text Compression with Denoising Transformers and Entropy Coding Compressive Transformers for Long-Range Sequence Modelling

Reference 51

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T21:07:24.272384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T19:54:39.953302Z digest=sha256:9fc26174bf0eea028483b2fb6e19f82826d344fd50c841d61e8366dcdf10f98a

Observation 541552d7-ebef-494b-87a6-2a9d41a105a0 · inbound

End-to-End Context Compression at Scale cites this paper.

End-to-End Context Compression at Scale Compressive Transformers for Long-Range Sequence Modelling

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-07-03T01:17:31.561208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T16:36:54.699174Z digest=sha256:0bd873bbd8178dcf4b068062614d060dc418a1c1aa7cf4edaf63dca38bdda860

Observation f97bef45-5deb-462e-b04d-0dce00056a4b · inbound

Dustin: Draft-Augmented Sparse Verification for Efficient Long-Context Generation with Speculative Decoding cites this paper.

Dustin: Draft-Augmented Sparse Verification for Efficient Long-Context Generation with Speculative Decoding Compressive Transformers for Long-Range Sequence Modelling

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T16:49:58.000560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T00:11:58.636939Z digest=sha256:6e92855cdb8f34a9e8f2e31573deb76e1afa378b67274fd4a4b6fb4c5aed816b

Observation 1249abb2-eec2-42cd-9c83-e77952abfa42 · inbound

Story Operators: Decomposing the Original $\to$ Sequel Transformation in Embedding Space cites this paper.

Story Operators: Decomposing the Original $\to$ Sequel Transformation in Embedding Space Compressive Transformers for Long-Range Sequence Modelling

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-04T19:30:07.243777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-25T21:19:44.766586Z digest=sha256:216e558c3d0b673f1406946dc4ad81f323968026244edd035b02080ebb163396

Observation e1d11da3-9332-44dc-8e91-60747613b0dc · inbound

Memory-Managed Long-Context Attention: Bounded Editable Memory with a Hard Lifecycle and Calibrated Sparse Fallback cites this paper.

Memory-Managed Long-Context Attention: Bounded Editable Memory with a Hard Lifecycle and Calibrated Sparse Fallback Compressive Transformers for Long-Range Sequence Modelling

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-06-30T09:54:35.249847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T09:46:57.030469Z digest=sha256:7f2b11fbcf81cf493beb8c10c43303fda2751b1e4140131e3a8c32d504ea203e

Observation eec98df6-c555-4703-bfae-9e3de2402520 · inbound

Memory-Managed Long-Context Attention: Bounded Editable Memory with a Hard Lifecycle and Calibrated Sparse Fallback cites this paper.

Memory-Managed Long-Context Attention: Bounded Editable Memory with a Hard Lifecycle and Calibrated Sparse Fallback Compressive Transformers for Long-Range Sequence Modelling

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-13T07:16:45.194991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T07:16:45.194991Z digest=sha256:a34f66dd77c83c3e4a5eaf13efa4aeb74b986fed0ab4d2b6991557df98c12f78

Observation 9a1f1709-3f85-4165-b378-81f4fdb07f70 · inbound

Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE cites this paper.

Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE Compressive Transformers for Long-Range Sequence Modelling

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-07-10T20:07:33.496053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-10T19:59:46.713277Z digest=sha256:00aabe1ef2ed17cbf57a3b2169dfcb5d689bddbadbe54a91a547c61a4f783c12

Observation d3270bb5-2f27-48cf-87d0-73d0420d46db · inbound

Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE cites this paper.

Jet-Long: Efficient Long-Context Extension with Dynamic Bifocal RoPE Compressive Transformers for Long-Range Sequence Modelling

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-13T06:47:09.927626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T06:47:09.927626Z digest=sha256:5d2b0d4fa7d0dfbdd4019c7dcc9d24e8527eb2f68bfc5a5a45079612d5b79685

Observation 50ada0fb-8be7-40f3-8d98-bae125d8ddbc · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents Compressive Transformers for Long-Range Sequence Modelling

Reference 97

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T01:36:44.161415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:6962757feebe0442a7ae1d68023c8f07c3e92936acdf3b1ad89b79e0e4a4f11b

Observation 55d8fb98-b1bc-4dbf-a4dc-258e23a394e2 · inbound

Structured Thoughts For Improved Reasoning And Context Pruning cites this paper.

Structured Thoughts For Improved Reasoning And Context Pruning Compressive Transformers for Long-Range Sequence Modelling

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-14T12:08:05.502310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T12:08:05.502310Z digest=sha256:75e9b1ac82eb0e2b8dc9271ffb5abceb09843d84d1a2cf9981b07598621232c0

Observation 10ea30ad-5d09-43c0-bb88-fa02bd9dc8f9 · inbound

UTS at ELOQUENT 2026 Voight-Kampff: structural shifts in AI writing bypass state-of-the-art detectors cites this paper.

UTS at ELOQUENT 2026 Voight-Kampff: structural shifts in AI writing bypass state-of-the-art detectors Compressive Transformers for Long-Range Sequence Modelling

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-02T04:51:21.716911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:51:21.716911Z digest=sha256:b44cf0bca3603a0bcc4c048d870d943339c546f0459ee8941da2de4987fd179f

Observation 40a5614f-3ab9-493e-9345-6b25e5a45a65 · inbound

High-accuracy Low-Bit KV-Cache Quantization via Local Distribution Restoration cites this paper.

High-accuracy Low-Bit KV-Cache Quantization via Local Distribution Restoration Compressive Transformers for Long-Range Sequence Modelling

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-02T09:50:33.078440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:50:33.078440Z digest=sha256:6c292512023828875d057b1ef91dfee64d3088d4e4a2ee4012f2df6c7def36b0

Observation c35e796e-91dc-4215-867d-4cc8591c3827 · inbound

SALT: Salience-Aware Lexical Trie for Long-Context Compression cites this paper.

SALT: Salience-Aware Lexical Trie for Long-Context Compression Compressive Transformers for Long-Range Sequence Modelling

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T17:53:16.638682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:53:16.638682Z digest=sha256:a22f2d1c9010da4792164e61bc4943d0b2358b457b9d394fe48fd2eaca008953

Observation 976465b4-3d89-4d2a-bd95-09dea427e9f3 · inbound

Windowed-MTP: Removing the Full-Context Draft-KV Tax at Million-Token Context cites this paper.

Windowed-MTP: Removing the Full-Context Draft-KV Tax at Million-Token Context Compressive Transformers for Long-Range Sequence Modelling

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T07:13:10.122353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:13:10.122353Z digest=sha256:0d571ab06548d261f920f2cb21f6e77993a159606f8d28de5d7e2999c148da20

Observation 6cad2ed6-b857-4284-aee9-8f1b4fb670f1 · inbound

MemSFT: Mitigating Alignment Tax with an External Parametric Memory cites this paper.

MemSFT: Mitigating Alignment Tax with an External Parametric Memory Compressive Transformers for Long-Range Sequence Modelling

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-01T01:59:49.824935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:59:49.824935Z digest=sha256:05920b8d0dac9b3f02557d232cfa36b4519d31ecaef44a8289c45d20f87f7dcd

Observation 7fcdedf1-afc9-4a53-8d29-2ed1fc60cea6 · inbound

Understanding Is Done Early: A Depth Division of Labor in Large Language Models and Its Use for Unbounded-Context Memory cites this paper.

Understanding Is Done Early: A Depth Division of Labor in Large Language Models and Its Use for Unbounded-Context Memory Compressive Transformers for Long-Range Sequence Modelling

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-07-31T12:56:45.774990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T12:56:45.774990Z digest=sha256:b4ff6cca60088c5d40e7393f30da657151c36e23969a3069878594a75c6e1ad1

Observation 12359230-a0c2-4073-873b-dd0e93983145 · inbound

A Taxonomy of Cognitive Capability Gaps in Generative and Agentic AI cites this paper.

A Taxonomy of Cognitive Capability Gaps in Generative and Agentic AI Compressive Transformers for Long-Range Sequence Modelling

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T04:52:58.464228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:52:58.464228Z digest=sha256:70d9296e5296cabb1532bf81d97c8ac756aeb997f3142601e64e2268b3095313