Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:44:34.878071Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 3 inbound Pith citation observations for arXiv:2506.07406.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:44:34.878071Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T22:20:52.947124Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
24 of 24 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 7de31c2a-3748-41af-83e5-97b57e0084c2 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Understanding intermediate layers using linear classifier probes
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 884b5900-8b90-48f0-9d6c-44d7f2449931 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Sparse Autoencoders Find Highly Interpretable Features in Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e292e361-3379-4728-9338-230cabb9bbb4 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models In-Context Learning Creates Task Vectors
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e71ec9b-eeab-4355-9b5f-4fca8bccb89f · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models RAVEL: Evaluating Interpretability Methods on Disentangling Language Model Representations
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99ca24ad-20bb-4290-94dc-230070a0d686 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96bf3527-f6c5-476c-89b8-a478013b0911 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Understanding deep image representations by inverting them.2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7465a8b7-834b-4ccf-a356-ba4da15f6817 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Multifaceted Feature Visualization: Uncovering the Different Types of Features Learned By Each Neuron in Deep Neural Networks
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f4976f0-6dab-4d11-b9ec-2840a7e3155b · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models The Linear Representation Hypothesis and the Geometry of Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a525b8f-a385-4852-9f8b-4e0c8b325bc5 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Gemma 2: Improving Open Language Models at a Practical Size
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f3ca20a-8fa8-4ff4-817c-02d1df676dd5 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2b18c6f-035b-4e61-bfb7-531a1aa1829f · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 841875b3-b3fd-4295-abc1-cd2f2616e098 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models the indirect object name in the prompt
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b2bb17f5-66e4-40da-9e8f-24620bb3cd3e · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Training is performed for 100,000 steps with a batch size of 2048 using the AdamW optimizer
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2e9583d3-d141-45cc-a038-8bc837b5cbc8 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models [City] is a city in the country of
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 767522fa-2384-4c3c-9976-8dde24a8ac93 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Then, [B] and [A] went to the [PLACE]. [B] gave a [OBJECT] to
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 46645982-f0df-4415-955e-f13c988cb566 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Scaling and evaluating sparse autoencoders
Reference 2009
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54d2460e-b0c6-4b0b-a2e0-f976a85526ea · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Mechanistic Interpretability for AI Safety -- A Review
Reference 2013
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 441ed5b6-b3e9-411b-81aa-d2ea158f8d60 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Towards Principled Evaluations of Sparse Autoencoders for Interpretability and Control
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cc61af8-21d7-4fe6-abc9-ff6dbf5e32d0 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Alexander Pan, Lijie Chen, and Jacob Steinhardt
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47020588-307b-4faa-a301-657d98bd9649 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Open Problems in Mechanistic Interpretability
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 989dcab7-d626-48ea-9f7b-930feae4c211 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models SelfIE: Self-Interpretation of Large Language Model Embeddings
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15a5bb06-6583-43b2-9dbb-38e809088d8f · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 905a8033-9350-41ff-9817-2c76ad7eb2e4 · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2a069249-8b0a-4bd1-96b1-37b4a6ef079c · outbound
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models Transformer Feed-Forward Layers Build Predictions by Promoting Concepts in the Vocabulary Space
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb51bfd6-9f2e-41b7-b344-ebd22b0de95f · inbound
Rep2Text: Decoding Full Text from a Single LLM Token Representation InverseScope: Scalable Activation Inversion for Interpreting Large Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6fc72081-ca3a-4f5c-bb10-c4f8d34b83fb · inbound
Mechanistic Interpretability of Cognitive Complexity in LLMs via Linear Probing using Bloom's Taxonomy InverseScope: Scalable Activation Inversion for Interpreting Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d512d461-2c4e-43ce-bc65-3833159c71c0 · inbound
PRISM: Recovering Instruction Sets from Language Model Activations InverseScope: Scalable Activation Inversion for Interpreting Large Language Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.