Pith. sign in

Paper Citation Record · LEDGER

Successor Heads: Recurring, Interpretable Attention Heads In The Wild

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2312.09230.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.09230 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T19:10:15.460977Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T07:06:44.476515Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 46b404c0-1b8e-4257-9633-5cfae82e5e95 · inbound

Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models cites this paper.

Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models Successor Heads: Recurring, Interpretable Attention Heads In The Wild

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-13T13:15:10.721813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-13T13:15:10.632115Z digest=sha256:3707d819023ec145700c666821bc9062bdb87c42b8db038a16d291dde44fe3b1

Observation 7d517ebe-3836-443d-a94b-263b5eaa08af · inbound

Understanding Multimodal LLMs: the Mechanistic Interpretability of Llava in Visual Question Answering cites this paper.

Understanding Multimodal LLMs: the Mechanistic Interpretability of Llava in Visual Question Answering Successor Heads: Recurring, Interpretable Attention Heads In The Wild

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T19:10:15.460977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:10:15.460977Z digest=sha256:a35a7a1aba100fa15701ebff829db9972feed3e0d253c6a3f01feeb2252981e1

Observation a0ba1185-ab48-45e3-9583-0794cbcd0c7e · inbound

ElastiFormer: Learned Redundancy Reduction in Transformer via Self-Distillation cites this paper.

ElastiFormer: Learned Redundancy Reduction in Transformer via Self-Distillation Successor Heads: Recurring, Interpretable Attention Heads In The Wild

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T14:38:51.842513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:38:51.842513Z digest=sha256:0549c51eed7665d1d11402025fccd3220ebe87e1ee8a464f8f0f35c0e2a87f50

Observation 6db12390-48a0-49f9-a2d5-f2cf4d168b88 · inbound

Structure Development in List-Sorting Transformers cites this paper.

Structure Development in List-Sorting Transformers Successor Heads: Recurring, Interpretable Attention Heads In The Wild

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T23:32:04.082858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T23:32:04.082858Z digest=sha256:121f9c1227acbb208ad27cb8efa44c9ff65483efc2aaf1ec60223d45a4c37fbd

Observation 56032f7f-3a9f-4bf5-8f60-5a264f06f9f6 · inbound

GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation cites this paper.

GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation Successor Heads: Recurring, Interpretable Attention Heads In The Wild

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T17:31:59.801158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:31:59.801158Z digest=sha256:bfa8156d64ea208e46e4d28abb89b9a1399639b282fe965d4ab9b78474e61044

Observation 28b19002-e07b-47b1-9546-a01bdb3afaee · inbound

On Mechanistic Circuits for Extractive Question-Answering cites this paper.

On Mechanistic Circuits for Extractive Question-Answering Successor Heads: Recurring, Interpretable Attention Heads In The Wild

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T11:01:28.865958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:01:28.865958Z digest=sha256:5638b2121c61ffaa9fb498b8d02351f35a53816ccb416817046c37bfe2f3d0ac

Observation ff19b961-2c6c-4c75-a2a8-7de80ebf4151 · inbound

To trust or not to trust: Attention-based Trust Management for LLM Multi-Agent Systems cites this paper.

To trust or not to trust: Attention-based Trust Management for LLM Multi-Agent Systems Successor Heads: Recurring, Interpretable Attention Heads In The Wild

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:32:17.224568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-19T11:30:47.877793Z digest=sha256:2983201377ecf04779c39970cd206842b9653d408a8391d6fffa964705550024

Observation f77a6226-41ca-4b14-be6c-f912467957cc · inbound

Modular Arithmetic: Language Models Solve Math Digit by Digit cites this paper.

Modular Arithmetic: Language Models Solve Math Digit by Digit Successor Heads: Recurring, Interpretable Attention Heads In The Wild

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T05:02:52.940577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:02:52.940577Z digest=sha256:2f63393d6327df797a7ff7954c96c7fb7a1264f2dc9af171869e8ab5a9b4034b

Observation 8a1ef151-094b-49c3-b881-e3e8c7075baa · inbound

HunyuanVideo-Foley: Multimodal Diffusion with Representation Alignment for High-Fidelity Foley Audio Generation cites this paper.

HunyuanVideo-Foley: Multimodal Diffusion with Representation Alignment for High-Fidelity Foley Audio Generation Successor Heads: Recurring, Interpretable Attention Heads In The Wild

Reference 2010

Resolution
unresolved
no resolver link, observed 2026-08-05T17:16:33.052383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:16:33.052383Z digest=sha256:6a58fd140427c375fc4d8488d348f80beef99fb0e2b49d887ea39eaf962f51c7

Observation bc5c4d12-15df-4b71-9025-a1d550f1ed7e · inbound

Model Science: getting serious about verification, explanation and control of AI systems cites this paper.

Model Science: getting serious about verification, explanation and control of AI systems Successor Heads: Recurring, Interpretable Attention Heads In The Wild

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T15:17:08.581962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:17:08.581962Z digest=sha256:084b137a0eccb6d1c2f65aea1f11556d135b2a70f501635506b616c99c7b18ff

Observation 62583b8d-f0b8-4d7a-9556-e81f18d18646 · inbound

Pattern Selectivity is Not Task-Causal Structure: A Cross-Architecture Mechanistic Study of Composed-Task Circuits in 1B-Class Language Models cites this paper.

Pattern Selectivity is Not Task-Causal Structure: A Cross-Architecture Mechanistic Study of Composed-Task Circuits in 1B-Class Language Models Successor Heads: Recurring, Interpretable Attention Heads In The Wild

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:06:44.478312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-28T07:07:58.170765Z digest=sha256:3eab1943c2cbe9b63afe367ec14cf2a8e64e4330d2a27285fac7c7216cd7477c