Pith. sign in

Paper Citation Record · LEDGER

Codebook Features: Sparse and Discrete Interpretability for Neural Networks

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2310.17230.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.17230 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:28:21.895444Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T10:15:44.367147Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ff968efa-cfb6-4539-9d2f-8b52d3b3b1e7 · inbound

Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know? cites this paper.

Reasoning about Uncertainty: Do Reasoning Models Know When They Don't Know? Codebook Features: Sparse and Discrete Interpretability for Neural Networks

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:21.895444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:28:21.895444Z digest=sha256:4ce50e2a3ba13d3b5babb65a34274174ee5819b64356ae534dc75f2be73e1942

Observation 9c0496bf-71fc-42ea-a76c-b215fc6f7781 · inbound

TopK Language Models cites this paper.

TopK Language Models Codebook Features: Sparse and Discrete Interpretability for Neural Networks

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T22:31:38.680344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:31:38.680344Z digest=sha256:1efa5c7e556e8310b40363961019f9601820cef23baca98c102dd74042ab318d

Observation e6bc741e-30f0-4536-8181-a1c612ea749d · inbound

Model Directions, Not Words: Mechanistic Topic Models Using Sparse Autoencoders cites this paper.

Model Directions, Not Words: Mechanistic Topic Models Using Sparse Autoencoders Codebook Features: Sparse and Discrete Interpretability for Neural Networks

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T11:01:42.904487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:01:42.904487Z digest=sha256:8920269491d5b29376de446d4d537924bb7a7951f70d3f2c5049ffc723e14c9e

Observation bc3db6eb-d65f-4b81-bb3c-8579b8948b65 · inbound

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet cites this paper.

Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet Codebook Features: Sparse and Discrete Interpretability for Neural Networks

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.492629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T07:50:04.813379Z digest=sha256:2656bff08e9554611af7c0f3905ad57757236f7a5687fd72071488cd24b7e915

Observation b7ff5da0-0933-4a50-b054-3657b3fa3cd2 · inbound

Explicit Fuzzy Logic in the Feed-Forward Layer: Self-Forgetting Quantifiers Discover Legible Grammatical-Licensing Detectors cites this paper.

Explicit Fuzzy Logic in the Feed-Forward Layer: Self-Forgetting Quantifiers Discover Legible Grammatical-Licensing Detectors Codebook Features: Sparse and Discrete Interpretability for Neural Networks

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:15:44.368706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T05:43:46.153089Z digest=sha256:c398d8fc8863f4ec87583066016d20ce2eeb77001f0bb7eef34fd68eb4a12688

Observation de236efa-f568-4a98-b507-9504b95eb3c0 · inbound

Legible-by-Construction: Attention and End-to-End Transformers cites this paper.

Legible-by-Construction: Attention and End-to-End Transformers Codebook Features: Sparse and Discrete Interpretability for Neural Networks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-11T20:06:55.102341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T20:06:55.102341Z digest=sha256:8d97288edde75bd795d61193f397a764597e4a66e5bd73c02e7075a121eaee65

Observation b9bd76ff-1d13-41cf-941b-e18a5592954b · inbound

Train the Model, Not the Reader: Decodability Supervision for Verifiable Activation Explanations cites this paper.

Train the Model, Not the Reader: Decodability Supervision for Verifiable Activation Explanations Codebook Features: Sparse and Discrete Interpretability for Neural Networks

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T10:03:02.404522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T10:03:02.404522Z digest=sha256:ba8b1edfe132aa8d44b000eb5a413ffeef0a76f9fa96e4510a4f3da37f7f5768