Pith. sign in

Paper Citation Record · LEDGER

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks

As of 16 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2505.12268.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.12268 v2

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:41:57.857103Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a5082a35-1917-4b11-9bee-9eba6da766fa · outbound

This paper cites What does bert look at? an analysis of bert’s attention.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks What does bert look at? an analysis of bert’s attention

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:41:58.720514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:41:57.697855Z digest=sha256:81f01a78e377303fa42f0e2371e11e2c5d09b5d91ef53e99af528e8b27a9f358

Observation 31e54323-2bdd-4f38-8adf-3d2c38cb5fb2 · outbound

This paper cites Sparse Autoencoders Find Highly Interpretable Features in Language Models.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Sparse Autoencoders Find Highly Interpretable Features in Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.713019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.713019Z digest=sha256:fdd715bfb7ab6c0383370ce9ebc7de0aa9bba9aa9be94f2518089ff1af053bfe

Observation 2758c388-4eea-4f35-bccd-c27306f92bc6 · outbound

This paper cites Scaling and evaluating sparse autoencoders.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Scaling and evaluating sparse autoencoders

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.735467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.735467Z digest=sha256:52361f554c766a41ee20840aa2d75779578d1cf9f0dfa22b4014a4e7359647db

Observation 40a80de9-8fb3-432b-8930-313ba78f6b85 · outbound

This paper cites Automatically Identifying Local and Global Circuits with Linear Computation Graphs.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Automatically Identifying Local and Global Circuits with Linear Computation Graphs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.746262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.746262Z digest=sha256:e581e1cbd2367f3faaffa37d408d179a2fac60072ca8be8c35959a9dcb758122

Observation 97258de6-1dd7-4a72-bca0-e374b93edbb0 · outbound

This paper cites A structural probe for finding syntax in word representations.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks A structural probe for finding syntax in word representations

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:41:58.658711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:41:57.756663Z digest=sha256:8a811254c6141ccaeb605806caf5610530c0f794dd3e34026ab2d32d2ae0b59b

Observation 8b18c6d1-e4d3-41ad-83a0-82e88774668c · outbound

This paper cites Language Models Use Trigonometry to Do Addition.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Language Models Use Trigonometry to Do Addition

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.767445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.767445Z digest=sha256:5165efad0904ea422951c97e4c2f953a13cfdbc959621f943ff18bb76fc87e52

Observation 1ffdf4d0-d5bb-4961-ac75-ea6d1f1ead9f · outbound

This paper cites Onboard deep lossless and near-lossless predictive coding of hyperspectral images with line-based attention.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Onboard deep lossless and near-lossless predictive coding of hyperspectral images with line-based attention

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T20:41:58.347588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:41:57.774976Z digest=sha256:ca7562effd15da19f94ce47a7b97fc502d484bfad5c7bf7189a2a7e35d0f624f

Observation 440a097c-656c-4f0f-a9a0-aa3c490c71ba · outbound

This paper cites Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.783451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.783451Z digest=sha256:715127a2e28a59a4b9a262f5f9d2bde041923cb9a8e80245cd0068d99c1ea67b

Observation 51320bff-1c9a-453c-981e-44cee6b1c88d · outbound

This paper cites A Mechanistic Interpretation of Arithmetic Reasoning in Language Models using Causal Mediation Analysis.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks A Mechanistic Interpretation of Arithmetic Reasoning in Language Models using Causal Mediation Analysis

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.806134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.806134Z digest=sha256:f86d218e55d6e73920bdf86d6ea2cb146986b11db2bd79389375fe90eec450ad

Observation d8059930-f293-4912-8a85-531183a2883b · outbound

This paper cites Planning in a recurrent neural network that plays Sokoban.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Planning in a recurrent neural network that plays Sokoban

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.812580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.812580Z digest=sha256:64db00cbe26e5669d7bd9ed3b3d0bda425a6677c4204571db826018297322832

Observation a874f307-8349-43ce-af75-229ecd86002e · outbound

This paper cites Greedy SLIM: A SLIM-Based Approach For Preference Elicitation.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Greedy SLIM: A SLIM-Based Approach For Preference Elicitation

Reference 17

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T20:41:58.180758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:41:57.822374Z digest=sha256:5ede15c67be647d900b8849461b50d793d4eece85563e069b5b488ef704e9e3c

Observation 0a511791-e9b3-46f4-a480-65a6d0a544fc · outbound

This paper cites Do Large Language Models Latently Perform Multi-Hop Reasoning?.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Do Large Language Models Latently Perform Multi-Hop Reasoning?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.829613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.829613Z digest=sha256:8cbaecf94b5c843ac1492ef9c761007795638c005302f25cc1f5f381a59de6c4

Observation 0f918cea-6a9c-42de-bfe7-3316fd7c7214 · outbound

This paper cites Back Attention: Understanding and Enhancing Multi-Hop Reasoning in Large Language Models.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Back Attention: Understanding and Enhancing Multi-Hop Reasoning in Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.837990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.837990Z digest=sha256:dc653c0e294533fccf4bf050d6064c7d244e93c121c5440ae93af5b28a81554c

Observation 0cd8144b-9469-465b-9e17-ea312c87ba26 · outbound

This paper cites The Same But Different: Structural Similarities and Differences in Multilingual Language Modeling.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks The Same But Different: Structural Similarities and Differences in Multilingual Language Modeling

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.847407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.847407Z digest=sha256:d03697fc45e73b0b748b9a68857728269dfee125b56b0db7fd03d06e1ebdd917

Observation 23e470d9-494f-4ca0-847e-47ff2cb4e4a5 · outbound

This paper cites Pre-trained Large Language Models Use Fourier Features to Compute Addition.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Pre-trained Large Language Models Use Fourier Features to Compute Addition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.857103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.857103Z digest=sha256:e682a0f7b6c0dd30f2c945bfbbc59d344921474780d4be977b8682da50d807ad

Observation 80d84e79-42d6-4a6c-9b71-27d15d99acab · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Training Verifiers to Solve Math Word Problems

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.707058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.707058Z digest=sha256:3112748ed80663bb017e2fbd636c3cf231de8b27efb9546a2395e9e7e10c29b9

Observation 5942f11e-34e8-4f12-9e3c-73ebb11e926f · outbound

This paper cites Andrew Stolfo, Atticus Geiger, David Friedman, and Vivek Srikumar.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Andrew Stolfo, Atticus Geiger, David Friedman, and Vivek Srikumar

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.798063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.798063Z digest=sha256:2f8e2cf806904a3a32569147844d62bf9e80fb7ba4cd2123b47da9ee41c11160

Observation dbb3afd0-cdf1-4716-93ec-40fce452e4c0 · outbound

This paper cites Lucy Gao, Lachlan Reynolds, Neel Nanda, and Chris Olah.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Lucy Gao, Lachlan Reynolds, Neel Nanda, and Chris Olah

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:41:58.685822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T20:41:57.723708Z digest=sha256:172d14bc3b6483e7da437c2b1583a6c17013b49e80127c59b84080fa2303d6d3

Observation 831539e5-aef9-43d1-9f5a-1fa53e94a015 · outbound

This paper cites Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.678412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.678412Z digest=sha256:2b77f733e315f862539f2918b7c3b50f0c89daa6fdade764accadc4dcdc6645b

Observation 050025f9-5eb3-4283-b2e6-1356f1c0b97d · outbound

This paper cites Hopping Too Late: Exploring the Limitations of Large Language Models on Multi-Hop Queries.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Hopping Too Late: Exploring the Limitations of Large Language Models on Multi-Hop Queries

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.669378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.669378Z digest=sha256:e7c8c6321c4c5c6eb083457a2046cb0ef58e9752ca42b92be07d784620bd050d

Observation 051ef9ac-a2d4-4b48-a5ac-1d69ebca18c1 · outbound

This paper cites Large Language Models Share Representations of Latent Grammatical Concepts Across Typologically Diverse Languages.

$K$-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks Large Language Models Share Representations of Latent Grammatical Concepts Across Typologically Diverse Languages

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:57.686649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:41:57.686649Z digest=sha256:50a9a51dcb962febb41650c918cb5abee6f8a77b3437dfc57cfd642c1a0d3e07

Pith citing papers

No inbound Pith citation observations are available.