Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2303.02536.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T14:55:02.030913Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
9
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation a2fc06f0-e7a3-46f3-afb6-47d45d3628ec · inbound
Localizing Model Behavior with Path Patching Finding Alignments Between Interpretable Causal Variables and Distributed Neural Representations
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8488324c-11e8-4be5-9705-375d81cb4875 · inbound
Towards Best Practices of Activation Patching in Language Models: Metrics and Methods Finding Alignments Between Interpretable Causal Variables and Distributed Neural Representations
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 024c1d07-87f5-4b2e-9bc5-4ecb81d25b8b · inbound
Interpretability in Parameter Space: Minimizing Mechanistic Description Length with Attribution-based Parameter Decomposition Finding Alignments Between Interpretable Causal Variables and Distributed Neural Representations
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 401c55ea-6fe8-4735-bc5c-fac8ed930e13 · inbound
What is a Number, That a Large Language Model May Know It? Finding Alignments Between Interpretable Causal Variables and Distributed Neural Representations
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68bbd057-2773-4fa5-a82f-ded588030d21 · inbound
Activation Reward Models for Few-Shot Model Alignment Finding Alignments Between Interpretable Causal Variables and Distributed Neural Representations
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a1e138e-84c0-4cc9-bea9-3ed10f1220bd · inbound
Local Linearity of LLMs Enables Activation Steering via Model-Based Linear Optimal Control Finding Alignments Between Interpretable Causal Variables and Distributed Neural Representations
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 816d3d25-ddd9-4c28-bbf9-050b72c14ae2 · inbound
Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs Finding Alignments Between Interpretable Causal Variables and Distributed Neural Representations
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ca0342c1-706c-46aa-b8d3-f0f377628d56 · inbound
Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes Finding Alignments Between Interpretable Causal Variables and Distributed Neural Representations
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3ca91b54-05de-4028-90bd-1be1c889cbd0 · inbound
From Mechanistic to Compositional Interpretability Finding Alignments Between Interpretable Causal Variables and Distributed Neural Representations
Reference 148
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation dc656d93-f134-406a-bc8d-c3e8ffb42b30 · inbound
Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Finding Alignments Between Interpretable Causal Variables and Distributed Neural Representations
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 66f400a5-dc44-4b54-8dde-ab0f3ec44b19 · inbound
Persistent Sparse Autoencoders: Learning Feature Timescales in Language Models Finding Alignments Between Interpretable Causal Variables and Distributed Neural Representations
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d75ec05-917e-428b-b71f-ce40c9a25fc6 · inbound
Emergent Misalignment Recruits a Pre-existing Persona Subspace Finding Alignments Between Interpretable Causal Variables and Distributed Neural Representations
Reference 132
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d4552c1-0b11-4981-8e23-faf7b4de599b · inbound
LAWFUL: Law-Aligned Witness for Faithful Use of Latents Finding Alignments Between Interpretable Causal Variables and Distributed Neural Representations
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.