Pith. sign in

Paper Citation Record · LEDGER

Parallelizing Linear Transformers with the Delta Rule over Sequence Length

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 39 inbound Pith citation observations for arXiv:2406.06484.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.06484 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 39 of 39 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T12:25:42.751276Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T01:36:44.246124Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 45cbfaf9-f7d9-45b4-8537-f62bbf4cdf09 · inbound

Learning to (Learn at Test Time): RNNs with Expressive Hidden States cites this paper.

Learning to (Learn at Test Time): RNNs with Expressive Hidden States Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 83

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T05:20:12.362233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T05:20:12.134340Z digest=sha256:6b4341046ce817e4a0fa2a58994b406ac93ca0cfccad0c792339fc9676d4da0b

Observation a2b0d8cd-b9c5-450a-8bba-bde9b09a046c · inbound

LASP-2: Rethinking Sequence Parallelism for Linear Attention and Its Hybrid cites this paper.

LASP-2: Rethinking Sequence Parallelism for Linear Attention and Its Hybrid Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T12:25:42.751276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:25:42.751276Z digest=sha256:9bdb8f930a1bd22ceddfa60fdbf708e9a40736997a172ffa6a291f1e2622abab

Observation d4861c86-2a08-49fc-a7d7-aec6e08425eb · inbound

An Uncertainty Principle for Linear Recurrent Neural Networks cites this paper.

An Uncertainty Principle for Linear Recurrent Neural Networks Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T22:10:15.942223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T22:10:15.942223Z digest=sha256:72c45947ab0d41f5f36db9f4eaba49dc8111630d5a8b2303ff52fc7950625358

Observation 0ec81779-3758-4fce-8c0d-3256b2e0dd59 · inbound

Solving Empirical Bayes via Transformers cites this paper.

Solving Empirical Bayes via Transformers Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T20:21:58.415098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:21:58.415098Z digest=sha256:9e56c36cc65daa06ca1181853ba1f88ec8f6f1b01c5c63a1ee7095e544f94044

Observation 18f01554-4540-4ce8-be3f-b9fbe424a5b2 · inbound

ModRWKV: Transformer Multimodality in Linear Time cites this paper.

ModRWKV: Transformer Multimodality in Linear Time Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:35:36.402647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:35:36.402647Z digest=sha256:3dd38336a8fbb2f73b46493cdf9fa59652987a040b9a0a9dafde64f71a2b5af6

Observation 02c654cb-d479-431e-b9f7-197bf653f528 · inbound

Understanding Transformer from the Perspective of Associative Memory cites this paper.

Understanding Transformer from the Perspective of Associative Memory Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:19:54.017185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:19:54.017185Z digest=sha256:96880cd6af095c0028a4df0a8dc8a2e5d45157dad0bb79bf6633711465768b08

Observation 23b43d5e-128a-4cf9-b784-405c7b5b86e2 · inbound

HAD: Hybrid Architecture Distillation Outperforms Teacher in Genomic Sequence Modeling cites this paper.

HAD: Hybrid Architecture Distillation Outperforms Teacher in Genomic Sequence Modeling Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:43.404040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:43.404040Z digest=sha256:1c7f0e8324090f0a02b3d426546de8c44778392f4a783df10a8da25cae5df2e9

Observation 232cce74-e538-4a71-8cb9-6ad92e78a480 · inbound

Scaling Reasoning without Attention cites this paper.

Scaling Reasoning without Attention Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:13:28.091568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:13:28.091568Z digest=sha256:25c05fd896d8fd348d782a47b270ff44dbf74287dd5583246f069723ed86f492

Observation 62fd2402-dff1-48e8-b4f3-c9e6fdceaf46 · inbound

Test-Time Training Done Right cites this paper.

Test-Time Training Done Right Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T11:25:45.215790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T11:25:45.153563Z digest=sha256:f1c0e31e9fdafba60a1e0cb7cacce23ef59ae85a2cd0213a23e242fa380c5a0c

Observation 83eca67f-71cf-45bc-b181-ec21efb3df9d · inbound

Cartridges: Lightweight and general-purpose long context representations via self-study cites this paper.

Cartridges: Lightweight and general-purpose long context representations via self-study Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T06:04:35.156790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:04:35.156790Z digest=sha256:1d40478b007511ee6f3d0c2e68f3516e95f2cc912f932f5f012ebf861e152e49

Observation ba368b4c-83f6-4f2c-9b86-76d1891b91af · inbound

SeerAttention-R: Sparse Attention Adaptation for Long Reasoning cites this paper.

SeerAttention-R: Sparse Attention Adaptation for Long Reasoning Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T05:06:32.732274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:06:32.732274Z digest=sha256:3189d5064febe9226b8129874447f0be3eff48c00e89de69285125b8eacc43b6

Observation 97f09e7b-fda8-42f7-ab1c-af5db940abdd · inbound

pLSTM: parallelizable Linear Source Transition Mark networks cites this paper.

pLSTM: parallelizable Linear Source Transition Mark networks Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T01:09:03.309496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:09:03.309496Z digest=sha256:2b128f0e4378fe8c720c182a438c7b04df591a5da5c3b7b03b2ab567407c311d

Observation 1e1eccb2-dc13-48c8-b255-08e6364afbe2 · inbound

TPTT: Transforming Pretrained Transformers into Titans cites this paper.

TPTT: Transforming Pretrained Transformers into Titans Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:48.547568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:48.547568Z digest=sha256:9d732fbf0eabfa02dde4a8d99cf9a47f9e1b4d411161054a84e814a623bae31c

Observation e8f220fa-a809-4837-a241-12a35269490b · inbound

A Survey on Latent Reasoning cites this paper.

A Survey on Latent Reasoning Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 123

Resolution
unresolved
no resolver link, observed 2026-08-06T19:14:32.936567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:14:32.936567Z digest=sha256:08cd40ec4efdeb6e402906b1ee74a41228e771fdabfe4c4eefed39e26f066f82

Observation f7501579-84d1-44ef-9bf6-57230956cc11 · inbound

Elucidating the Design Space of Decay in Linear Attention cites this paper.

Elucidating the Design Space of Decay in Linear Attention Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T05:29:21.681430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:29:21.681430Z digest=sha256:3a4405106a4577d2cc7c618107ebfa9c6db4afb49939b0b19205c13c256a0963

Observation 7356987f-d11c-4a9d-b54c-fcb96e8ae67a · inbound

A Vision Toward Energy-Efficient Domain-Specific Artificial Intelligence Models and Agents cites this paper.

A Vision Toward Energy-Efficient Domain-Specific Artificial Intelligence Models and Agents Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 161

Resolution
unresolved
no resolver link, observed 2026-08-04T08:12:26.022787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:12:26.022787Z digest=sha256:7691743af5a29b8c65dc0224329d802df911a717cd63f6a1fcbd7ce85c46b049

Observation 2c7c4f8b-9510-468a-90ef-3938ee3967c0 · inbound

Nirvana: A Specialized Generalist Model With Task-Aware Memory Mechanism cites this paper.

Nirvana: A Specialized Generalist Model With Task-Aware Memory Mechanism Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:05:48.075690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T03:05:06.069642Z digest=sha256:097cef37ca4e313a140c4b3a951856b31af0a9f942d8ed6f0b3aad094a0dc90f

Observation 2ff72c95-29fb-42f6-bd33-9d57a928ab5a · inbound

Kimi Linear: An Expressive, Efficient Attention Architecture cites this paper.

Kimi Linear: An Expressive, Efficient Attention Architecture Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 115

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:49:10.947759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T23:49:10.555255Z digest=sha256:46d5fd8acd3555c2ed23ba64137f562b462677a970a79dbf52d895d4669399d3

Observation 89921422-127d-466e-a755-21bce9b43d24 · inbound

LADY: Linear Attention for Autonomous Driving Efficiency without Transformers cites this paper.

LADY: Linear Attention for Autonomous Driving Efficiency without Transformers Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T06:42:11.959356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:42:11.959356Z digest=sha256:17a9e502535ab98f3b63912e48abc8101021b7304c0bb7fa38bc3aededdc1254

Observation d9ea9905-e22c-49f2-91f1-fd79b33813c1 · inbound

Physics of Language Models: Part 4.1, Architecture Design and the Magic of Canon Layers cites this paper.

Physics of Language Models: Part 4.1, Architecture Design and the Magic of Canon Layers Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-03T15:22:55.477441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:22:55.477441Z digest=sha256:5e039141d6dea6891507a8df8abb3b59f70cff36a30392c9eeb7f52b33d616be

Observation 0d74fb44-14e2-4840-82e4-fa5e16a48c4a · inbound

ZipMap: Linear-Time Stateful 3D Reconstruction via Test-Time Training cites this paper.

ZipMap: Linear-Time Stateful 3D Reconstruction via Test-Time Training Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:26:17.502452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T16:21:16.229770Z digest=sha256:96affc1412ce5906bf4e66220adb945bdba9fc1f696df6b6e5aecebe8008ff1a

Observation 6056f5d8-f335-4de9-8225-71322553850c · inbound

Phase-Associative Memory: Sequence Modeling in Complex Hilbert Space cites this paper.

Phase-Associative Memory: Sequence Modeling in Complex Hilbert Space Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:40:48.610341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:39:50.311059Z digest=sha256:f7463889e53e615a9cd00f0da81e18bd8ab4fd1287243fc4963d5ee8b1b217a1

Observation 5767fc1c-ee17-4eb9-94b1-5624f8e8337b · inbound

In-Place Test-Time Training cites this paper.

In-Place Test-Time Training Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:49.017551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:07:47.174513Z digest=sha256:26138eca991575c7e03924539922b9d1448d1f8981afefb6ed9335eb738865ef

Observation cefb2429-bd4c-4845-84e1-d1a29bca075d · inbound

COREY: Entropy-Guided Runtime Chunk Scheduling for Selective Scan Kernels cites this paper.

COREY: Entropy-Guided Runtime Chunk Scheduling for Selective Scan Kernels Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:06:03.419188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:10:02.445509Z digest=sha256:68a58fdc3e5d34ccede011be2ac51a7ab86f529060eedd25cdbc6bf281ad0b35

Observation 13e9f8fe-4a18-4c48-a4d5-589d3698fa85 · inbound

Adaptive Memory Decay for Log-Linear Attention cites this paper.

Adaptive Memory Decay for Log-Linear Attention Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:50:56.408116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T01:02:56.848785Z digest=sha256:fe762535e771c52c8716317ecdd474dc7854e2d899324ecf27eee564c6cc9315

Observation 92437f3f-a79a-45a7-b2b5-9048abe9d5a1 · inbound

WriteSAE: Sparse Autoencoders for Recurrent State cites this paper.

WriteSAE: Sparse Autoencoders for Recurrent State Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:59:28.639301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-14T20:53:40.666929Z digest=sha256:ce503ab8e59f20621ecaf8dfa69e0a4ad48791c0856ed28939ac0489573acb03

Observation 5f2fda31-0cc1-4eb1-b180-e08fd22bd39b · inbound

WriteSAE: Sparse Autoencoders for Recurrent State cites this paper.

WriteSAE: Sparse Autoencoders for Recurrent State Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T04:59:45.229904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T04:59:11.877068Z digest=sha256:8b6a253f5af49fb0ef47b770d8903042aedd7d57038c6cbf551d3d36ed2907e9

Observation acf0bbbd-8876-4b89-8c2b-64d76c440ad0 · inbound

Towards Understanding Self-Pretraining for Sequence Classification cites this paper.

Towards Understanding Self-Pretraining for Sequence Classification Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:33:58.947730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-21T05:29:58.809024Z digest=sha256:8e49235e58337b26e99372c0e00f3171eb44915a74cd1124614c6df83ab8af7b

Observation 6b38ad3b-883b-47a6-9e75-d3ab1ffdac2c · inbound

Pretraining Recurrent Networks without Recurrence cites this paper.

Pretraining Recurrent Networks without Recurrence Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 140

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:26:56.635118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T02:09:01.018909Z digest=sha256:e0abb0f2ff27eb8d332ad7a4a40c09c8cbfcaec6a716a1b0743dda8833be6535

Observation 0bc8bce5-eb20-43ed-907b-64dc34790fea · inbound

Pretraining Recurrent Networks without Recurrence cites this paper.

Pretraining Recurrent Networks without Recurrence Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 137

Resolution
unresolved
no resolver link, observed 2026-08-02T12:20:56.082616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:20:56.082616Z digest=sha256:3cdcddc439ac24adc855b5b7825c6cb1c9487523f728a796947779736b37ca30

Observation ae2b1780-dc47-47ba-b973-84e95c2a6006 · inbound

Reversible Foundations: Training a 120B Sparse MoE through State-Preserving Scaling cites this paper.

Reversible Foundations: Training a 120B Sparse MoE through State-Preserving Scaling Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:17:08.659060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T22:55:09.477413Z digest=sha256:0d05ff01bf0ed21ccd84e1887915b5eb47e65b2fcd44c5835386343cb1e3b407

Observation b240309e-de36-4b54-a789-cbaff4debd74 · inbound

Q-Delta: Beyond Key-Value Associative State Evolution cites this paper.

Q-Delta: Beyond Key-Value Associative State Evolution Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 91

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T23:07:26.542463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T18:30:51.523567Z digest=sha256:5f854cadd159441036eeef48456661cea0792cf8a769616e7bd172d81a46c053

Observation 19a33f76-fce7-4f0f-a010-4f98ee69db51 · inbound

UltraQuant: 4-bit KV Caching for Context-Heavy Agents cites this paper.

UltraQuant: 4-bit KV Caching for Context-Heavy Agents Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:39:30.395444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T17:49:02.835708Z digest=sha256:83532811ff4fdacc0ffb449a656b3bc4e11721da735671bcc0f4ca25aa76a670

Observation 9ca11e76-f7f5-48d7-b5e9-c377b9e88a02 · inbound

ELiTeFormer: An Efficient Transformer for FPGAs cites this paper.

ELiTeFormer: An Efficient Transformer for FPGAs Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 62

Resolution
unresolved
no resolver link, observed 2026-07-12T00:55:09.690245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:55:09.690245Z digest=sha256:3b86fa49d695d32bafd6bb83491de63d5be5a702ec956f5804e933496469632d

Observation 43c01be1-5d43-46ce-bc59-8222c393ed08 · inbound

Sparse Delta Memory: Scaling the State of Linear RNNs through Sparsity cites this paper.

Sparse Delta Memory: Scaling the State of Linear RNNs through Sparsity Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 114

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T12:46:14.646570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-09T12:40:09.036905Z digest=sha256:beecbfd697ab9fc863056aa46463bb49d442300f4a93c0fb865885998904db85

Observation 4301c229-8a9b-405e-b51f-8591ae6b2b84 · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 140

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:36:44.247325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:74d59e5eb462bf3df926648b1d1c9403a7992fdf477ed892dfd63c6a6c529e79

Observation 9568a843-40ac-4db2-9004-f5155f3a2d5b · inbound

The Orthogonalized Read Is a Removable Training Scaffold for Recurrent Memory cites this paper.

The Orthogonalized Read Is a Removable Training Scaffold for Recurrent Memory Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T09:06:11.171935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:06:11.171935Z digest=sha256:a4d10948afd3951b7879fafcdf6707921bd50f90e0654966f53c01cf78ac44b7

Observation dfb0e7d7-7f23-40be-a02f-8603824b40bc · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T13:43:51.979801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:43:51.979801Z digest=sha256:b07bb8f5fd1b0c71adc182300f14b2d4fc8ee73c284e9d557de40a08e6009cdd

Observation 47293eaf-457b-4b0a-be73-ed5dfe5943e0 · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs Parallelizing Linear Transformers with the Delta Rule over Sequence Length

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:11:48.805634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:11:48.805634Z digest=sha256:758eb3c6a54f6f93083ca40f421a9f371bde5f8e3b2f352b8cff8986653c53cf