Pith. sign in

Paper Citation Record · LEDGER

Finding Neurons in a Haystack: Case Studies with Sparse Probing

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 67 inbound Pith citation observations for arXiv:2305.01610.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.01610 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 67 of 67 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:47:49.650382Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T15:06:18.074949Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 623dd1e3-363a-4150-9427-50d88a0e2707 · inbound

Towards Best Practices of Activation Patching in Language Models: Metrics and Methods cites this paper.

Towards Best Practices of Activation Patching in Language Models: Metrics and Methods Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T11:56:11.133944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-17T11:56:11.053897Z digest=sha256:6104595ed329aea97defb8a12b4ac39d6dcd8d48543643b85c23e20ae323c2f9

Observation 66e64a2d-093d-43fb-8247-ebdf1f5d0454 · inbound

Scaling and evaluating sparse autoencoders cites this paper.

Scaling and evaluating sparse autoencoders Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T17:47:23.168842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-12T17:47:23.089288Z digest=sha256:31e7ab37777ac218659b32b174b8cc6b2593e97fe4a26ea8da6e7c59d95e93de

Observation 5498284f-de55-4794-a395-45b3bcdcea7e · inbound

PERFT: Parameter-Efficient Routed Fine-Tuning for Mixture-of-Expert Model cites this paper.

PERFT: Parameter-Efficient Routed Fine-Tuning for Mixture-of-Expert Model Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T21:56:03.304209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:56:03.304209Z digest=sha256:bb458c81aa56cc2e892e05a60ef4e1a632d3dcb3879b8a3f158083c99750bfd4

Observation c259ab0b-95d4-49bd-ada5-10e226902946 · inbound

Understanding Multimodal LLMs: the Mechanistic Interpretability of Llava in Visual Question Answering cites this paper.

Understanding Multimodal LLMs: the Mechanistic Interpretability of Llava in Visual Question Answering Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T19:10:15.465197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:10:15.465197Z digest=sha256:bb934eab5767013b55dd27c2b423101dd85f04de0ae4ae72f6f46bc87d7fe59c

Observation 7db41ce0-0b5e-444d-8225-18c00a5e21e4 · inbound

A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions cites this paper.

A Survey on Uncertainty Quantification of Large Language Models: Taxonomy, Open Research Challenges, and Future Directions Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T20:37:54.747005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:37:54.747005Z digest=sha256:01897c7b88345d3e6115232856642e2f629343ec6c1722c68ce20370e9e19989

Observation a1ced7f3-3595-45ef-a0da-162188957faa · inbound

GPT-2 Through the Lens of Vector Symbolic Architectures cites this paper.

GPT-2 Through the Lens of Vector Symbolic Architectures Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T18:26:31.785201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:26:31.785201Z digest=sha256:ecb7abe687935ac8e73abfd45ece14d0b1b86d94cf5fed3b63253d325f75e395

Observation 5d20f218-8d4f-499c-b6cf-ff1317b83b93 · inbound

Joint Knowledge Editing for Information Enrichment and Probability Promotion cites this paper.

Joint Knowledge Editing for Information Enrichment and Probability Promotion Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T10:18:44.147493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T10:18:44.147493Z digest=sha256:e1bc3adfc5765306e735538785dbba95479dc34c0d049f3822b1b6452fd6fcdd

Observation 6a2186de-82a7-4bd0-af6d-2cd0cf24cbdd · inbound

Flash Interpretability: Decoding Specialised Feature Neurons in Large Language Models with the LM-Head cites this paper.

Flash Interpretability: Decoding Specialised Feature Neurons in Large Language Models with the LM-Head Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T22:12:09.831365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:12:09.831365Z digest=sha256:a284aff0c730717cedfcd9e4f5241b27f145964fdbed33538f38db87a4491ecb

Observation 2f2cd971-2d58-43c8-9e55-cc1d6327c0e9 · inbound

Rethinking Evaluation of Sparse Autoencoders through the Representation of Polysemous Words cites this paper.

Rethinking Evaluation of Sparse Autoencoders through the Representation of Polysemous Words Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T21:28:15.897390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:28:15.897390Z digest=sha256:7e4961c24864ec3fc0354a414ff5193aed812d04de6949943763ea940249c7d2

Observation 8ebfe189-4b58-4945-9c96-e72c608cf0cc · inbound

Transcoders Beat Sparse Autoencoders for Interpretability cites this paper.

Transcoders Beat Sparse Autoencoders for Interpretability Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T22:24:53.852861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:24:53.852861Z digest=sha256:5e2aa769c2a5092ca70e6b7ab7dd73b122302437ef077902c9b8d23e795b8d6d

Observation e35e1049-ac4d-4c1a-8430-9702ebe0ec5b · inbound

Partially Rewriting a Transformer in Natural Language cites this paper.

Partially Rewriting a Transformer in Natural Language Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-09T22:21:50.745427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:21:50.745427Z digest=sha256:15df6cad4b82f9fff720a958d975bb3e8ba22e9bdb0d8e9fe7e4f1761739091d

Observation 103cac41-3917-4ee7-a161-33950e5bb43e · inbound

What is a Number, That a Large Language Model May Know It? cites this paper.

What is a Number, That a Large Language Model May Know It? Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T15:05:03.732462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:05:03.732462Z digest=sha256:0864419d9451130e01d09130fb06fee92e570b79d7183d61aa73160426d1923c

Observation 3ddc6f00-e32d-4913-a754-bef4387f3807 · inbound

Discovering Chunks in Neural Embeddings for Interpretability cites this paper.

Discovering Chunks in Neural Embeddings for Interpretability Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T14:29:19.995790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:29:19.995790Z digest=sha256:b7c00b9cb7453b603a115e81682d169ef283da86f9e61187819191908f6b2380

Observation 58f73ddc-4c8b-418f-9f63-4dff5d1a44e1 · inbound

Neurons Speak in Ranges: Breaking Free from Discrete Neuronal Attribution cites this paper.

Neurons Speak in Ranges: Breaking Free from Discrete Neuronal Attribution Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:12:30.701871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-23T04:09:40.210836Z digest=sha256:5321c122bcc24f0518379c277958551dec1bc4e6e97677718f41a55d9ed515c0

Observation 0ac42879-be90-4943-8fe1-63e4d9fa8328 · inbound

Beyond Cross-Modal Alignment: Measuring and Leveraging Modality Gap in Vision-Language Models cites this paper.

Beyond Cross-Modal Alignment: Measuring and Leveraging Modality Gap in Vision-Language Models Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T02:55:19.761554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-23T02:53:34.960236Z digest=sha256:cc2a9da8b9d77f03a35e344325520a2f07e15a0bade0281b84ffcbc10644ee4e

Observation 09ad25b0-e162-47c1-924b-08c936c2e61a · inbound

Disentangling Linguistic Features with Dimension-Wise Analysis of Vector Embeddings cites this paper.

Disentangling Linguistic Features with Dimension-Wise Analysis of Vector Embeddings Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T11:47:49.650382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:47:49.650382Z digest=sha256:39125888327418b13b68e4e543ca17dd7396e38609995adf7add61e40a2dc22d

Observation e605bb39-eece-4860-8457-abdf04276ec8 · inbound

Representation Learning on a Random Lattice cites this paper.

Representation Learning on a Random Lattice Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T05:41:19.082242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:41:19.082242Z digest=sha256:96746e686e6b1093d4ca84ea1b404f6edf6de91af36920c6f22e2d1bc736d884

Observation 9a574b75-8ac1-4083-b4d5-c7cd70aff09b · inbound

Investigating task-specific prompts and sparse autoencoders for activation monitoring cites this paper.

Investigating task-specific prompts and sparse autoencoders for activation monitoring Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:12.185035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:38:12.185035Z digest=sha256:258aea4f278dbf2dda92292e6ed63474eeb37d9b45d7a9b885fa24a7f74e2a5c

Observation a0bbd9e0-7a43-4066-89d8-4f5e374e5c7a · inbound

Polysemy of Synthetic Neurons Towards a New Type of Explanatory Categorical Vector Spaces cites this paper.

Polysemy of Synthetic Neurons Towards a New Type of Explanatory Categorical Vector Spaces Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-16T05:04:37.709731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:04:37.709731Z digest=sha256:a218b0bc1183081ac859883133e16cf0c3f37a2cf93b2ea01f6c0a8df2d47df6

Observation 08284514-5534-4b73-99dc-445649be6fe6 · inbound

Recovering Event Probabilities from Large Language Model Embeddings via Axiomatic Constraints cites this paper.

Recovering Event Probabilities from Large Language Model Embeddings via Axiomatic Constraints Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T22:41:50.785870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:41:50.785870Z digest=sha256:2fd21cf36889f5a2cc72600f3db6b6186df0fe722ddad2705554d0775946f1b6

Observation 41431003-4af2-4675-912a-b5fa5f3a99f9 · inbound

Emergent Specialization: Rare Token Neurons in Language Models cites this paper.

Emergent Specialization: Rare Token Neurons in Language Models Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T20:30:15.290065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:30:15.290065Z digest=sha256:fccc36a56a4a7f94d22c18e4ca7b1684fdf4991c86e41ac3a99f030b4da161ae

Observation e2b6e765-9c95-40d3-b93b-fc305eb639bd · inbound

Void in Language Models cites this paper.

Void in Language Models Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:37:00.997871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:37:00.997871Z digest=sha256:689a4aa5bf6871e06ce4eb82b698911d2437e86c7f66949c5fc431c13198bc7c

Observation 0b93f34b-3004-4596-9677-e10cdc1fc9de · inbound

Understanding Gated Neurons in Transformers from Their Input-Output Functionality cites this paper.

Understanding Gated Neurons in Transformers from Their Input-Output Functionality Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:37.783968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:42:37.783968Z digest=sha256:1231f6db592f5a972b19bccccac1796cf2dde6f77bf0e65536cfdd576f82b85e

Observation 6deba5fe-bdc3-4692-b2bc-1b2846e5a728 · inbound

ALPS: Attention Localization and Pruning Strategy for Efficient Alignment of Large Language Models cites this paper.

ALPS: Attention Localization and Pruning Strategy for Efficient Alignment of Large Language Models Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:29.711577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:31:29.711577Z digest=sha256:b893c556de7059c71c7fd82ea6606cd548206dcfdd8dafcd90662127754d4bc5

Observation 362b2aa4-dbb3-4c37-a61c-e3c8ff3dbc3c · inbound

Paths Not Taken: Understanding and Mending the Multilingual Factual Recall Pipeline cites this paper.

Paths Not Taken: Understanding and Mending the Multilingual Factual Recall Pipeline Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:58:17.838066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:58:17.838066Z digest=sha256:cf31ba8fd3752c052e3aeee4aa497993379dab745a2fb21431959fcf72be2456

Observation d0767621-9cc3-40ec-b899-b19887780b2e · inbound

Pretrained LLMs Learn Multiple Types of Uncertainty cites this paper.

Pretrained LLMs Learn Multiple Types of Uncertainty Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:44:20.352171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:44:20.352171Z digest=sha256:d68d8bbf47b429caf842df6c63144a43718df85c0911eec4379543a0c70e398f

Observation 7c3aee70-7e65-4b93-aa95-46e43127fa63 · inbound

Understanding the learned look-ahead behavior of chess neural networks cites this paper.

Understanding the learned look-ahead behavior of chess neural networks Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:47.434967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:47.434967Z digest=sha256:639bc07c99c71e3ffb81df07a0d75db014d49ae9aaa4a0ef0adb39b2c24c4eb3

Observation 39745a64-56d9-4af9-b3fb-c26887384e68 · inbound

Distinct Computations Emerge From Compositional Curricula in In-Context Learning cites this paper.

Distinct Computations Emerge From Compositional Curricula in In-Context Learning Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:53.737338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:53.737338Z digest=sha256:e8ada0719fc7a4fd5c2b405afa0a4f2a7d1f75ca8e66038fc0f0e7066608f288

Observation c4a128f1-6e34-488c-bc27-d9c2535c4cd6 · inbound

The Compositional Architecture of Regret in Large Language Models cites this paper.

The Compositional Architecture of Regret in Large Language Models Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T23:59:50.905197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:59:50.905197Z digest=sha256:fcb648e40a77fbec0d0f735a72be04af93ce6978579e2793cf23f0e84e62630c

Observation 599725ac-c13c-40aa-b525-a83d1e2d9e05 · inbound

Cross-Layer Discrete Concept Discovery for Interpreting Language Models cites this paper.

Cross-Layer Discrete Concept Discovery for Interpreting Language Models Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:03:04.944808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:03:04.944808Z digest=sha256:607800fa90ab3e65037a2909676b7a0af2742d34aa7f5b8ff570689003b7a4fc

Observation 3d8a695e-fcd4-4c45-8d73-08ec6490e71e · inbound

Learning to Skip the Middle Layers of Transformers cites this paper.

Learning to Skip the Middle Layers of Transformers Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T22:40:54.224975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:40:54.224975Z digest=sha256:cd94312fb111044c7579180d2f316c5936b48217dbd9f8fb8a32f15486120801

Observation f7a901bc-7b60-4059-94c4-2b8109e29806 · inbound

Scaling laws for activation steering with Llama 2 models and refusal mechanisms cites this paper.

Scaling laws for activation steering with Llama 2 models and refusal mechanisms Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:05:03.916915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:05:03.916915Z digest=sha256:70762caa8eea2b2ea06c77029398d75128ec5dbf9987e67a9d691d4d3d95853d

Observation f25a916e-e558-4824-8691-95b94162b92c · inbound

On the transferability of Sparse Autoencoders for interpreting compressed models cites this paper.

On the transferability of Sparse Autoencoders for interpreting compressed models Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:24:45.726080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:24:45.726080Z digest=sha256:a8509adf142d59f913bb4a853a1d50d285459d363573883f4f34ab9953ed857c

Observation 10667aba-16f3-4bbb-8fb3-6066f4d26e0a · inbound

Adaptive Lattice-based Motion Planning cites this paper.

Adaptive Lattice-based Motion Planning Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T17:43:04.889252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:43:04.889252Z digest=sha256:28849551c950d468b3ab733d60b4164104db3bc14b95bf80f371a2236a1afe04

Observation 6af67497-3285-4607-a6b2-d10b926d5721 · inbound

NEAT: Concept driven Neuron Attribution in LLMs cites this paper.

NEAT: Concept driven Neuron Attribution in LLMs Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T18:01:19.625643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:01:19.625643Z digest=sha256:deb45e1eeefacad953d066ba745e374fedda8486fb9c46e24ec66a2560b124b5

Observation c5b71cd4-73d5-4923-b094-3d5c2d3cb41a · inbound

Distribution-Aware Feature Selection for SAEs cites this paper.

Distribution-Aware Feature Selection for SAEs Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-05T14:27:25.436740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:27:25.436740Z digest=sha256:31081d3bb912c8f202842adf782143004db8f3566ef39bec99bb2496fed16feb

Observation aa775b40-e6c4-4be5-8fde-2b1b6f6d5fb8 · inbound

Safe-SAIL: Towards a Fine-grained Safety Landscape of Large Language Models via Sparse Autoencoder Interpretation Framework cites this paper.

Safe-SAIL: Towards a Fine-grained Safety Landscape of Large Language Models via Sparse Autoencoder Interpretation Framework Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-18T18:16:43.758605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-18T18:13:01.662828Z digest=sha256:f736bb691568c2c249d8c14b2e44b188b6cf751c4f445fe38522ebf8e3ea1195

Observation c0a413f7-9c02-4123-bd2a-7837170e3caa · inbound

Towards Atoms of Large Language Models cites this paper.

Towards Atoms of Large Language Models Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T15:20:32.162715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T15:20:32.162715Z digest=sha256:9fb4a025e50966adbb12cec695a26920deb3c176825c1b49f479fbaed4d7add7

Observation 6650ebcd-6d81-464a-b994-179b62f0f833 · inbound

Foundation Models for Discovery and Exploration in Chemical Space cites this paper.

Foundation Models for Discovery and Exploration in Chemical Space Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 266

Resolution
verified exact
arxiv_id, observed 2026-05-18T05:52:24.900507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T05:52:10.848118Z digest=sha256:84b78b40b0c39496179ab4a30a580f922073b3591ea1a0f3d611dd8646731cc7

Observation 24ca7d6e-3630-41ec-a256-4183d1d8b8df · inbound

Friends and Grandmothers in Silico: Localizing Entity Cells in Language Models cites this paper.

Friends and Grandmothers in Silico: Localizing Entity Cells in Language Models Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T09:24:05.580140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T09:21:42.717424Z digest=sha256:feb8edc85fa57ab2c7562d9fae5c991fccd7849fe98ed6eca92c1eb10ed78f4b

Observation 636c7b7c-65f6-4208-b2de-2b57f449d10c · inbound

Do Audio-Visual Large Language Models Really See and Hear? cites this paper.

Do Audio-Visual Large Language Models Really See and Hear? Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:58:15.812312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T20:56:19.815569Z digest=sha256:14ab34cbf058bd26d3cac2d28030322576450debdc06de2e3e25be198cb5bbec

Observation 923101a2-e534-4f74-908b-2c6d47954703 · inbound

Darkness Visible: Reading the Exception Handler of a Language Model cites this paper.

Darkness Visible: Reading the Exception Handler of a Language Model Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:00:57.187345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T18:43:41.723816Z digest=sha256:791c4e69df973d8c51df20c0e0054682326eca7931691efaba1d29b4bf35eeef

Observation 94d03fc2-0ed1-43de-b5ed-f7ed7293b01f · inbound

I Walk the Line: Examining the Role of Gestalt Continuity in Object Binding for Vision Transformers cites this paper.

I Walk the Line: Examining the Role of Gestalt Continuity in Object Binding for Vision Transformers Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T07:56:01.377444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T16:53:04.513790Z digest=sha256:528f01f551a9b0c7c0ce66098ba1168ade4d581983126c00097e5cde8ace434f

Observation 9a321963-bd03-4eb8-b056-9d11d2a70b87 · inbound

Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs cites this paper.

Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:21:00.007981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-10T16:41:52.440793Z digest=sha256:3b1b3b6e6833cc62c296b8cb8c219335e4ec1fe0104fcb3cbebe6e3315a0fb68

Observation 06cda7a6-565e-4f67-809e-289409ea56f2 · inbound

There Will Be a Scientific Theory of Deep Learning cites this paper.

There Will Be a Scientific Theory of Deep Learning Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 207

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:21:08.990987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-09T20:11:17.616190Z digest=sha256:ae8d4aab775a38fc1db2b94fcccc79a8451670e2f340018709139380d3bba975

Observation c1728dbb-7cbe-4ea9-8c3f-7d9314b5db88 · inbound

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs cites this paper.

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T10:01:27.604209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-07T08:34:14.310656Z digest=sha256:10da3991154afeba26cea29c186b7dcf5b5800e7f191cca962a7fae56c263f4a

Observation 5663786d-8c59-4609-bb92-df6293eb8867 · inbound

Tree SAE: Learning Hierarchical Feature Structures in Sparse Autoencoders cites this paper.

Tree SAE: Learning Hierarchical Feature Structures in Sparse Autoencoders Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T03:15:54.416400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T03:13:58.543525Z digest=sha256:b510a3fc8425abdfbe9fb95a8dd4ce1261ecfcd28af302ed97210e0125f5b8da

Observation 96d9e37f-d4c1-466c-9b40-db61efb156fd · inbound

Tree SAE: Learning Hierarchical Feature Structures in Sparse Autoencoders cites this paper.

Tree SAE: Learning Hierarchical Feature Structures in Sparse Autoencoders Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:16:25.410351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T03:35:50.776347Z digest=sha256:af98da47d721b00cf5511b3194a80ab1ef1cc906f231051f1ebb21d025a7165d

Observation 74ee3f79-bfe2-4e93-9288-3e3ab371ba4e · inbound

A Single Neuron Is Sufficient to Bypass Safety Alignment in Large Language Models cites this paper.

A Single Neuron Is Sufficient to Bypass Safety Alignment in Large Language Models Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:46:14.379445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-12T01:41:44.898358Z digest=sha256:70beaf11e366c8ffdf525d4ad99a660a88e2e5ecdfadd348cfb00370519e3d24

Observation f3a388ca-986a-462e-88f4-746f2fb6f97d · inbound

WriteSAE: Sparse Autoencoders for Recurrent State cites this paper.

WriteSAE: Sparse Autoencoders for Recurrent State Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 92

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T20:59:28.796625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-14T20:53:40.666929Z digest=sha256:24558ac914d063e5c5e71922c12cc2c59bec26b04baa70199564311c36f14032

Observation e931232f-2681-4130-855b-e1e77c050006 · inbound

WriteSAE: Sparse Autoencoders for Recurrent State cites this paper.

WriteSAE: Sparse Autoencoders for Recurrent State Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 92

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T04:59:45.227604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-15T04:59:11.877068Z digest=sha256:15aabd972a0d6d057cbdb9ac4cb36b4068b787fec8057dda43b2b40ca0d03658

Observation c575fa01-d9cc-454c-b481-8235eaf19aed · inbound

WriteSAE: Sparse Autoencoders for Recurrent State cites this paper.

WriteSAE: Sparse Autoencoders for Recurrent State Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 92

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T21:53:47.255923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-20T21:49:47.934339Z digest=sha256:cd0381fa15f7639b337d45e68a1b6adb34c4490e842b519430989e643a96b0a7

Observation aed69250-b958-44f4-9dc0-9a610b0b8fc7 · inbound

WriteSAE: Sparse Autoencoders for Recurrent State cites this paper.

WriteSAE: Sparse Autoencoders for Recurrent State Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:49:50.156950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T07:46:41.159688Z digest=sha256:bd7bdd63f4a4652699cf0c1c3c6b7aaebacab9ff5a0d12d509b1d9d9bd5d2bc9

Observation 4a50f9dc-3054-4091-9c2e-b1433370d0ea · inbound

Chessformer: A Unified Architecture for Chess Modeling cites this paper.

Chessformer: A Unified Architecture for Chess Modeling Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T12:33:17.151389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T12:28:55.936689Z digest=sha256:09b5e14adcbe8c8b484576e59d735a5b547779e0f4b839a4550b8f8e5011ff63

Observation dd8d8636-4908-4445-8cdc-7a13cff88130 · inbound

Mechanistic Interpretability for Learning Assurance of a Vision-Based Landing System cites this paper.

Mechanistic Interpretability for Learning Assurance of a Vision-Based Landing System Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T06:59:45.613970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T06:56:23.842767Z digest=sha256:89c7e421ea45eac3b1812920a7e0f7c3ee5fc4f3029e4c0117f70a8a34ae11db

Observation 07d64787-a15f-438a-89ff-1914881b1c47 · inbound

Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection cites this paper.

Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T14:03:29.899390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T13:53:27.306664Z digest=sha256:bae4171bbc437cbd950817655388d416fce4360c3cbb61f966c03876c49ff510

Observation 3be4657e-7245-4189-9fc4-c416c43d5ea1 · inbound

A Geometric View for Understanding Concept Learning and Neuron Interpretation in Sparse Autoencoders cites this paper.

A Geometric View for Understanding Concept Learning and Neuron Interpretation in Sparse Autoencoders Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:47:09.880393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-27T22:22:50.474397Z digest=sha256:eebedb8ef6b17b57885fd359ee48f0e30387c2aeb46715ce87f601da6d719ec7

Observation 3c7b6049-ae6d-4c68-acde-9e33d0c80609 · inbound

The Amplifying Mirror: Locating and Steering the Partisan Direction inside a Large Language Model cites this paper.

The Amplifying Mirror: Locating and Steering the Partisan Direction inside a Large Language Model Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T23:07:26.683370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T18:30:14.923773Z digest=sha256:8ed0391e73c9ad98641850afb9a9e73b45b5cf18135cf2188954824c2b251bcd

Observation 00bcaa89-c0dd-4856-9b0d-2f8e0d13231e · inbound

ICA Lens: Interpreting Language Models Without Training Another Dictionary cites this paper.

ICA Lens: Interpreting Language Models Without Training Another Dictionary Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T09:17:49.080628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T10:21:58.878499Z digest=sha256:1d04367f6ffb2b3c463ed1d4de9f80de8eed55c441240acfd33fbc9cb751019a

Observation 43bd97e9-8bd9-4a19-a489-39a40bc52ad4 · inbound

From Texts to Scores: Tracing the Emergence of Essay Quality Representations in Large Language Models cites this paper.

From Texts to Scores: Tracing the Emergence of Essay Quality Representations in Large Language Models Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:59:34.280364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-26T17:17:23.250463Z digest=sha256:c9e32fa69122a159803dbccd543a4b4bea18646b688f2db07c3d88bd151b7e11

Observation 8721b206-13da-4095-9e91-3c029648a611 · inbound

Critical Percolation as a Synthetic Data Model for Interpretability cites this paper.

Critical Percolation as a Synthetic Data Model for Interpretability Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:49:29.699180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T17:41:29.317167Z digest=sha256:4034730c6e76c4ecdaa0a3e7567ffbb4c13ca880f3d02d5c0f0cf007e6266581

Observation 862c9084-0db2-4bce-ad60-137bcc7398ed · inbound

Behavioral and Representational Evidence of Binomial Ordering Preferences in Large Language Models cites this paper.

Behavioral and Representational Evidence of Binomial Ordering Preferences in Large Language Models Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 256

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:39:37.875428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-26T14:18:11.215278Z digest=sha256:4b4ccbb5e29807e1e089128a16c2f9e8c2fe807dd8c4c5cd26f504ec09a7bca2

Observation 82de02b5-f4c0-4f02-972a-06cae982d7ee · inbound

Individual Parameters in Weight-Sparse Transformers Appear Interpretable cites this paper.

Individual Parameters in Weight-Sparse Transformers Appear Interpretable Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T05:48:11.148130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:48:11.148130Z digest=sha256:c3c92fae4494f3a53d70a3510b6f701ce74b8a3d4dc741c20b8abdb7f7101ac1

Observation f5b2eb1b-1fb1-44ad-950b-e4e237767793 · inbound

Mechanistic Interpretability for Neural Networks: Circuits, Sparse Features and Symbolic Reasoning cites this paper.

Mechanistic Interpretability for Neural Networks: Circuits, Sparse Features and Symbolic Reasoning Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T15:06:18.076196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T14:58:58.363330Z digest=sha256:0045e469f37964b195df34cb1d28fa7a0ccad4da28f02983b55762f4e0d1e64b

Observation a618b0e9-6ab8-47c6-96d7-5ce7574e489d · inbound

How are linear representations learned? Exact solutions to the dynamics of abstraction cites this paper.

How are linear representations learned? Exact solutions to the dynamics of abstraction Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T06:19:30.027337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T06:19:30.027337Z digest=sha256:da7b242a35ee65cd9cece40030c1421221c96741140934f602e96481c5250d9c

Observation d26fe9ab-a94d-4528-a4ed-9bb07f16a09d · inbound

U-Lens: Supporting User Uncertainty Management in Long-Form LLM Responses cites this paper.

U-Lens: Supporting User Uncertainty Management in Long-Form LLM Responses Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-14T10:32:54.448282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:32:54.448282Z digest=sha256:161eac5d5940825a4cc2279081b57efeeb1a5ba89122e685652a99acd5b28bb5

Observation b3e6d4cc-123d-46cc-857f-af84d30a0d3b · inbound

CADENCE: A Cardiac Atom Dictionary for Interpretable Neural Concept Extraction from ECG Foundation Models cites this paper.

CADENCE: A Cardiac Atom Dictionary for Interpretable Neural Concept Extraction from ECG Foundation Models Finding Neurons in a Haystack: Case Studies with Sparse Probing

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T03:03:56.744673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:03:56.744673Z digest=sha256:d5cb2345b9a3dc0db01b433544e55f7f61a3317fae97bcad4148042e0e9e3012