Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2104.07143.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:25:05.642287Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T12:46:57.379107Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation a175f726-d97d-4094-8d9e-db3489192e8d · inbound
Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small An Interpretability Illusion for BERT
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 30bde55e-1f39-443b-ae73-d4fad517bce7 · inbound
Localizing Model Behavior with Path Patching An Interpretability Illusion for BERT
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 54bdea91-580e-4f4f-8ca2-32e2a07e3f68 · inbound
Improving Dictionary Learning with Gated Sparse Autoencoders An Interpretability Illusion for BERT
Reference 202
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3cee2c1b-c56f-4e33-a67e-0590ee663b44 · inbound
Scaling and evaluating sparse autoencoders An Interpretability Illusion for BERT
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 5999d966-490f-4f43-b9f2-9921abce60fc · inbound
Towards Utilising a Range of Neural Activations for Comprehending Representational Associations An Interpretability Illusion for BERT
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be99f224-566d-4813-8519-f3251303824a · inbound
Incremental Sentence Processing Mechanisms in Autoregressive Transformer Language Models An Interpretability Illusion for BERT
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20d6f206-4a8f-4290-84b2-4b636a1693a5 · inbound
Inferring Functionality of Attention Heads from their Parameters An Interpretability Illusion for BERT
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eff1370b-756b-4a39-a1e1-b341cef9d4f2 · inbound
The Heap: A Contamination-Free Multilingual Code Dataset for Evaluating Large Language Models An Interpretability Illusion for BERT
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 777759fa-b0ae-43f2-a7fd-eb557a7d9d8e · inbound
Perspectives for Direct Interpretability in Multi-Agent Deep Reinforcement Learning An Interpretability Illusion for BERT
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87cc22af-9945-4683-b20e-ef73fe38ec6a · inbound
Evaluating Explanations: An Explanatory Virtues Framework for Mechanistic Interpretability -- The Strange Science Part I.ii An Interpretability Illusion for BERT
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a67009e-afc0-4b94-9e35-3417fcd244a6 · inbound
Recovering Event Probabilities from Large Language Model Embeddings via Axiomatic Constraints An Interpretability Illusion for BERT
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1b8db61-ae1a-4685-a0a9-8957651cb965 · inbound
Evaluating Neuron Explanations: A Unified Framework with Sanity Checks An Interpretability Illusion for BERT
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35896df0-11a0-4992-afe6-415d35e253b0 · inbound
Sign-Aware Gated Sparse Autoencoders: Modeling Anticorrelated Features with Bi-Jump-ReLU Activations An Interpretability Illusion for BERT
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f25ab80d-7a18-4e33-9c39-bcd4f3b98a1f · inbound
Sign-Aware Gated Sparse Autoencoders: Modeling Anticorrelated Features with Bi-Jump-ReLU Activations An Interpretability Illusion for BERT
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6f48bcd-4a4f-4411-b962-b0732e2f3383 · inbound
Many Circuits, One Mechanism: Input Variation and Evaluation Granularity in Circuit Discovery An Interpretability Illusion for BERT
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 677968aa-ccce-4f23-bd82-78289adf54e4 · inbound
The Entanglement Wall: Activation-Space Probes as Risk Detectors, Not Context Adjudicators An Interpretability Illusion for BERT
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03de5da2-0cbf-40a8-914d-212329928d9e · inbound
Where You Measure Decides What You Measure: Position Selection in Ablation-Based SAE Evaluation An Interpretability Illusion for BERT
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.