Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T12:32:45.226075Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2411.17182.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T12:32:45.226075Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
48 of 48 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8ac51476-fb4a-47f2-b0f9-cd2b04cc9ae8 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Repulsive attention: Rethinking multi-head attention as bayesian inference
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b35fff6e-3134-45ad-a3a2-ce6b93eed7b0 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Layer Normalization
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ea967a3-0fb2-4185-baf8-45ae8ea67187 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Spectrally-normalized margin bounds for neural networks
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5462d016-0753-4c0c-ae96-de693445e988 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Birth of a transformer: A memory viewpoint
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9578f20e-5e1e-4c5b-95dc-6e8750604d8b · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Attention approximates sparse distributed memory
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0090f6f0-0841-4359-b3ac-4e7327cbb323 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Invariant scattering convolution networks
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 3e9edc9b-69b6-42fd-ada6-314db41fb187 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Emerging properties in self-supervised vision transformers
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b9218ed-4289-47eb-bacd-3b04d66e3770 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Transformer interpretability beyond attention visualization
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c11ae912-1203-478c-847b-9370b9337f47 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Randaugment: Practical automated data augmentation with a reduced search space
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b8ff53b-013c-45ee-b6b2-276f5cba3f2e · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Analyzing transformers in embedding space
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f52c75f1-0776-4cc8-b6a4-d0f8c87c02cd · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models An image is worth 16x16 words: Transformers for image recognition at scale
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d1094ff-ed3a-49a6-8b29-fca4c410bafb · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models A mathematical framework for transformer circuits
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c06a0860-7503-4381-acf7-bd362f606f16 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models The emergence of clusters in self-attention dynamics
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e79ac14f-aa14-437b-bf24-6cd0f9b3ac1f · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Patchscopes: A unifying framework for inspecting hidden representations of language models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bf391b7c-adc1-4acb-b43b-61c758f08718 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Learning fast approximations of sparse coding
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 605f79d1-605a-4da1-a9f0-03c135299011 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models The Forward-Forward Algorithm: Some Preliminary Investigations
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1be9dafb-f19b-4cee-a6cc-ca8e4612b738 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Energy transformer
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1625818f-5370-4e2e-a8bb-73ec627ab404 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Fantastic generalization measures and where to find them
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7608b21f-bdd6-41c6-98c5-ab7723c5504e · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models A new measure of rank correlation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f2f13eb-9e2d-4757-8686-ed2a53edbc07 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models On large-batch training for deep learning: Generalization gap and sharp minima
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 250b883e-35af-43ae-9acd-9f8f564a1372 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Kingma and Jimmy Ba
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76d36dfa-a66a-4e74-ad0c-1ce875852f48 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Tracr: Compiled transformers as a laboratory for interpretability
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bc6f5695-7460-492c-8500-81222d17031c · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Omnigrok: Grokking beyond algorithmic data
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69c8da1a-a61c-4f39-855b-c2171387f5d5 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Segmentation of multivariate mixed data via lossy data coding and compression
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 501377e5-6383-49eb-a7cb-e64f16f231c7 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Pac-bayesian model averaging
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 73fe6ed6-92e0-4665-9c76-4021e94b5bc2 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Universal hopfield networks: A general framework for single-shot associative memory models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b50e8e18-e4a0-4c14-8af1-a2d7dc9b730e · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Algorithm unrolling: Interpretable, efficient deep learning for signal and image processing
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3344aab9-2dbb-4b06-8a69-8fcaf690904a · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Progress measures for grokking via mechanistic interpretability
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0584541d-c4ff-4f99-bb6e-cf0ad3df2db9 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Exploring generalization in deep learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d01d58ef-e403-45e8-bbeb-38bc23881973 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models A PAC-bayesian approach to spectrally-normalized margin bounds for neural networks
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a6007c5e-c231-47cd-8547-3df21ea368f2 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Path-sgd: Path-normalized optimization in deep neural networks
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a29f89a8-f7d3-4be4-a5aa-84cddea91a4d · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Norm-based capacity control in neural networks
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf19aabb-2092-4d15-a8ac-824a19b2ef4d · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Theoretical foundations of deep learning via sparse representations: A multilayer sparse model and its connection to convolutional neural networks
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 68370c3f-74ac-4fc1-865f-be6b2b4cf1b2 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Grokking: Generalization Beyond Overfitting on Small Algorithmic Datasets
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c16a147-3160-4d09-925c-5c70977a2500 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Hopfield networks is all you need
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d047be21-d60e-4e4d-96ac-a3cb81671714 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Unraveling attention via convex duality: Analysis and interpretations of vision transformers
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e8f5ac9e-88a5-44d7-a4e7-f40c3afbcb3f · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Biological learning in key-value memory networks
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7ceec1d0-12e0-4188-9409-b85ef544dc90 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models On the uniform convergence of relative frequencies of events to their probabilities
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45595754-205e-4650-8973-eb26e14d9a7c · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Attention is all you need
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54fbf4f8-9de5-4311-930a-8de8af52ca77 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Interpretability in the wild: a circuit for indirect object identification in GPT-2 small
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16708be7-7ec9-4462-b6f0-49fd059e8c47 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Thinking like transformers
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 28500a26-2d01-4722-86b7-8319da737fa4 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Graph neural networks inspired by classical iterative algorithms
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5819e8a2-a4ec-4038-be03-839a1c45ad48 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Transformers from an optimization perspective
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0b214848-862c-4a0d-8529-e072e6f554e2 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Attentionviz: A global view of transformer attention
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 12cf0273-156d-4d0c-b1cc-529692a0833f · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models White-box transformers via sparse rate reduction
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 048031ec-59eb-4fc9-a8a0-d30640be511d · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Learning diverse and discriminative representations via the principle of maximal coding rate reduction
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation af831bc5-e02c-4fd1-8cad-c1d2357de761 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Unveiling Transformers with LEGO: a synthetic reasoning task
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7346d03-3e0f-4e9f-9460-f178836751b5 · outbound
An In-depth Investigation of Sparse Rate Reduction in Transformer-like Models Unresolved cited work
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.