Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T00:02:46.864302Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 2 inbound Pith citation observations for arXiv:2602.11852.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T00:02:46.864302Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-13T17:43:52.716947Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-13T17:48:03.356591Z
34 of 34 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a1bfac4d-6b43-4614-85e4-264b835edc9d · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af9e2236-62b9-4057-90da-077a4e4f759e · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design in the”, “of the
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3da37f5-1912-488d-a78c-5bace7c0377c · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design why does my dog eat poo p?
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b5c5f2e-3e6c-48c2-92d8-1952187014cd · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Alignment faking in large language models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adedb359-c754-481d-8284-abb4d84791ec · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Mamba: Linear-Time Sequence Modeling with Selective State Spaces
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46b28b14-7707-4f1e-a7d1-9d6078015bd1 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfe80cb7-8e10-468a-9df2-844fa94f8921 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Interpreting Attention Layer Outputs with Sparse Autoencoders
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4374789-4366-41cb-9250-59ca772f2aa5 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Decoupled Weight Decay Regularization
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b0442a0-0e4a-43bc-b578-13a5d9bbdb8b · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Revisiting Small Batch Training for Deep Neural Networks
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fee98cb1-71ec-4910-927b-fcd4aa72dee5 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Pointer Sentinel Mixture Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fce6703-a533-4f72-8c27-92c5febd96ee · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design A practical review of mecha- nistic interpretability for transformer-based language models.arXiv preprint arXiv:2407.02646,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a63bffcf-1e9f-4821-adeb-6a6073a781d3 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Sentence-BERT: Sentence embeddings using Siamese BERT- networks
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc58d4da-ff62-462b-89c8-92d5539d6325 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Neural Machine Translation of Rare Words with Subword Units
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41941a36-071c-442f-a066-8c600b8b94ba · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design BERT Rediscovers the Classical NLP Pipeline
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb975022-1a6e-46f5-adac-0c72f8705585 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design LLaMA: Open and Efficient Foundation Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5148a92f-9c00-473b-864f-105e0c29e300 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08ba0b77-c02a-48c3-9a40-3b02ad92be0f · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ecbbd32-0886-4350-8b15-0a4a915caa04 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Qwen3 Technical Report
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 834de153-2f61-4a87-88ae-1c3586f40b90 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Image classification at supercomputer scale
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bfe7acc-6173-4475-b784-2f6042fd6c09 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Towards Best Practices of Activation Patching in Language Models: Metrics and Methods
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2572600-d266-46aa-9b73-c18d7d38e2fe · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Deconstructing What Makes a Good Optimizer for Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5bea32e-e11e-4c37-b82f-13aa38042664 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design to help us produce the prototype interpretability html
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eab81462-3c7d-4139-81fe-216efd394dbb · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design You are analyzing a single prototype (a neuron-like feature) from a neural language model.\n
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b00eba98-5974-4fa3-a4bd-be6ff0f33e8b · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Unresolved cited work
Reference 512
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40fa6dbd-3e7a-44e4-9b18-20d62b4e15f5 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Automatically Interpreting Millions of Features in Large Language Models
Reference 1995
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ced3996-6364-4a23-a77f-5bc065fea788 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design GLU Variants Improve Transformer
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 020f9809-d3ca-4704-8415-ab7576d37bca · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Jamba: A Hybrid Transformer-Mamba Language Model
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26f2a1fc-a6e3-4916-ab4e-87dbfafc3d56 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c653af77-502d-40b8-8e35-261f826ecdd2 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Toy Models of Superposition
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15f46023-078f-4a3e-ae84-15d5872d299b · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design ReZero is All You Need: Fast Convergence at Large Depth
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75b0f373-ec3b-4ecf-9551-f1139e62d36c · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design Dissecting Recall of Factual Associations in Auto-Regressive Language Models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c45c933-8f46-44dc-921d-f26437a88557 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design The Llama 3 Herd of Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51a80759-3718-4f16-ad06-a22a23c087d3 · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design What Does BERT Look At? An Analysis of BERT's Attention
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94ee83c8-71d3-4a93-938f-b8bb4da36bfc · outbound
Prototype Transformer: Towards Language Model Architectures Interpretable by Design A Transformer and Prototype-based Interpretable Model for Contextual Sarcasm Detection
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fda4066-2b23-4177-af82-33dbb4f77ae8 · inbound
Collapse-Free Prototype Readout Layer for Transformer Encoders Prototype Transformer: Towards Language Model Architectures Interpretable by Design
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f31719cd-a815-4e8c-a88d-662a3464dd50 · inbound
Graph Memory Transformer (GMT) Prototype Transformer: Towards Language Model Architectures Interpretable by Design
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.