Pith. sign in

Paper Citation Record · LEDGER

Knowledge Neurons in Pretrained Transformers

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 31 inbound Pith citation observations for arXiv:2104.08696.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.08696 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 31 of 31 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:37:56.902551Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

42
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fd415246-94a6-458e-98a3-be0a638c2aeb · inbound

GPT-NeoX-20B: An Open-Source Autoregressive Language Model cites this paper.

GPT-NeoX-20B: An Open-Source Autoregressive Language Model Knowledge Neurons in Pretrained Transformers

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-24T12:34:28.368782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-24T12:33:37.701655Z digest=sha256:9021f35d2f3bde3aa08ff257423f097b5ff37ea27a567b38de1c7610c5b0e0c2

Observation 1304339b-c5f8-4f6e-801c-1fb3df7bf0e5 · inbound

Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling cites this paper.

Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling Knowledge Neurons in Pretrained Transformers

Reference 146

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:45:17.741750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-15T17:45:17.540282Z digest=sha256:12bb3fea3fb0c3227f9ae828d9a23f320a9fb9ecc1e90a74d1d4e0e7d2658c80

Observation 26f81c94-0a26-4ad4-b8ae-e6875d100bac · inbound

SalUn: Empowering Machine Unlearning via Gradient-based Weight Saliency in Both Image Classification and Generation cites this paper.

SalUn: Empowering Machine Unlearning via Gradient-based Weight Saliency in Both Image Classification and Generation Knowledge Neurons in Pretrained Transformers

Reference 138

Resolution
verified exact
arxiv_id, observed 2026-05-16T17:56:23.415905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-16T17:56:23.281678Z digest=sha256:bef7607503acfc6dfd32be040d100ec15c598fa3175222c544c1979eec8cf8e4

Observation 1c06dc29-88dd-4d94-8401-5d1dae438d87 · inbound

Neural Incompatibility: The Unbridgeable Gap of Cross-Scale Parametric Knowledge Transfer in Large Language Models cites this paper.

Neural Incompatibility: The Unbridgeable Gap of Cross-Scale Parametric Knowledge Transfer in Large Language Models Knowledge Neurons in Pretrained Transformers

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:37:56.902551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:37:56.902551Z digest=sha256:b60a0bbfd6bee59fa1e03d20d9612c13a252cc08ed54a5fd8c46e8357495f9f9

Observation 471b2b56-b5ad-45cc-92e4-9259004b0f83 · inbound

Towards a Science of Causal Interpretability in Deep Learning for Software Engineering cites this paper.

Towards a Science of Causal Interpretability in Deep Learning for Software Engineering Knowledge Neurons in Pretrained Transformers

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:50.727326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:29:50.727326Z digest=sha256:7eb51608b669281057cf1a591a87a6300f9d303879cd0854e3c3f62b0be5f7d8

Observation 0daf8a15-2d6a-4b70-8464-3db14bdaaaf6 · inbound

Locate-then-Merge: Neuron-Level Parameter Fusion for Mitigating Catastrophic Forgetting in Multimodal LLMs cites this paper.

Locate-then-Merge: Neuron-Level Parameter Fusion for Mitigating Catastrophic Forgetting in Multimodal LLMs Knowledge Neurons in Pretrained Transformers

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:01:15.482314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:01:15.482314Z digest=sha256:99960810971eb6472eba345b6b86e3bdf8ae57d81bcf397f28fb006a7a01bff5

Observation fd8655f5-1f3f-47ba-b36d-2ec22f393a16 · inbound

Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models cites this paper.

Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models Knowledge Neurons in Pretrained Transformers

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:18.445140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:18.445140Z digest=sha256:20c807cb1f125bfedfd7ac3a11b15a7a017993b79df31208ebe49ca176c701b0

Observation b277f457-bf29-4123-a67d-9a56eab948d0 · inbound

TRACE for Tracking the Emergence of Semantic Representations in Transformers cites this paper.

TRACE for Tracking the Emergence of Semantic Representations in Transformers Knowledge Neurons in Pretrained Transformers

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:10.581513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:42:10.581513Z digest=sha256:4a1c6410bac171c2b53a8730fa39f553e02149ed95bb3ee9bd709f4bc2003cc4

Observation 4c6b2c78-02bb-4e00-b1b2-7c7666218dfa · inbound

A Graph Perspective to Probe Structural Patterns of Knowledge in Large Language Models cites this paper.

A Graph Perspective to Probe Structural Patterns of Knowledge in Large Language Models Knowledge Neurons in Pretrained Transformers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:14.587960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:21:14.587960Z digest=sha256:a708072e84ec6e9ec187ce2f02e3c4b2167d34f537bf5ca23d873b1c6f1e04de

Observation df639f4c-da81-48cd-8826-711e4cb44b8a · inbound

Unveiling Instruction-Specific Neurons & Experts: An Analytical Framework for LLM's Instruction-Following Capabilities cites this paper.

Unveiling Instruction-Specific Neurons & Experts: An Analytical Framework for LLM's Instruction-Following Capabilities Knowledge Neurons in Pretrained Transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:43:20.561182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:43:20.561182Z digest=sha256:16d8d261a4d9a199f9da2b69dc3224a2e32e40b8805f2b55dd6af3d14a0de318

Observation ef4f804c-3d72-48a3-bd26-cb370bd77d53 · inbound

SplitLoRA: Balancing Stability and Plasticity in Continual Learning Through Gradient Space Splitting cites this paper.

SplitLoRA: Balancing Stability and Plasticity in Continual Learning Through Gradient Space Splitting Knowledge Neurons in Pretrained Transformers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:15:55.841089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:15:55.841089Z digest=sha256:6061ba7d8f0596a2ef1be4adacb515673cbbbae10f40dc40ba93b6725f6d4f59

Observation fd85d045-d595-4567-9672-723e93756b34 · inbound

Rhetorical Text-to-Image Generation via Two-layer Diffusion Policy Optimization cites this paper.

Rhetorical Text-to-Image Generation via Two-layer Diffusion Policy Optimization Knowledge Neurons in Pretrained Transformers

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:04:51.455988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:04:51.455988Z digest=sha256:0a19452422cdd32505268a395fef299e1a37141ef8c965ac5981839e749b5cbe

Observation 537dc18c-1cbf-4940-84d7-b3f9abf6ee1e · inbound

Establishing Trustworthy LLM Evaluation via Shortcut Neuron Analysis cites this paper.

Establishing Trustworthy LLM Evaluation via Shortcut Neuron Analysis Knowledge Neurons in Pretrained Transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:54:26.227834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:54:26.227834Z digest=sha256:2c69ae966d0bbc17f8250984bdeabe0b092d1cb5da9164e119c7d3e8b68f7c95

Observation 5152826f-02a9-45ee-b7b4-2a6667c24a56 · inbound

Attribution-Guided Pruning for Insight and Control: Circuit Discovery and Targeted Correction in Small-scale LLMs cites this paper.

Attribution-Guided Pruning for Insight and Control: Circuit Discovery and Targeted Correction in Small-scale LLMs Knowledge Neurons in Pretrained Transformers

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:53:04.139080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T08:52:45.818050Z digest=sha256:449e8298ea19dbb116212b1726d1d2a3a1198539618872c67fd9a76bfc5e7054

Observation 5b74a649-33b4-4b3c-a370-b2692a9d0bf9 · inbound

QF: Quick Feedforward AI Model Training without Gradient Back Propagation cites this paper.

QF: Quick Feedforward AI Model Training without Gradient Back Propagation Knowledge Neurons in Pretrained Transformers

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T19:57:55.322574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:57:55.322574Z digest=sha256:bfaa35787a4b9ed6b0c4660636e83e2f4a5374fd75939b05af169d4cc193a087

Observation 49ff2f9e-cc47-4ab3-90bf-13a3c5f955f4 · inbound

A Survey on Proactive Defense Strategies Against Misinformation in Large Language Models cites this paper.

A Survey on Proactive Defense Strategies Against Misinformation in Large Language Models Knowledge Neurons in Pretrained Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:03:45.306645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:03:45.306645Z digest=sha256:90a937dc41d4035b8c5235bf8aafd94ebe624d06b9483f50333ceb729af400fb

Observation 8add6027-20fe-4941-adcf-8e0ad70d6728 · inbound

What Should LLMs Forget? Quantifying Personal Data in LLMs for Right-to-Be-Forgotten Requests cites this paper.

What Should LLMs Forget? Quantifying Personal Data in LLMs for Right-to-Be-Forgotten Requests Knowledge Neurons in Pretrained Transformers

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T17:21:33.568852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:21:33.568852Z digest=sha256:9ad7f117248d3491293e60e9f2907815d48ef97286c90ed01faa5d84bb417fa9

Observation 593048a1-e30d-43f1-a342-4f464c232644 · inbound

Granular Concept Circuits: Toward a Fine-Grained Circuit Discovery for Concept Representations cites this paper.

Granular Concept Circuits: Toward a Fine-Grained Circuit Discovery for Concept Representations Knowledge Neurons in Pretrained Transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T05:30:17.038384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:30:17.038384Z digest=sha256:ce6543342ee4be8e735618d3452d4ea2fb0dcf4e5e906eba4c1739098a30cf5b

Observation 44ec3d41-53c9-4d3c-8cf3-2c365e167fb9 · inbound

Unveiling the Response of Large Vision-Language Models to Visually Absent Tokens cites this paper.

Unveiling the Response of Large Vision-Language Models to Visually Absent Tokens Knowledge Neurons in Pretrained Transformers

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T11:15:19.454823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:15:19.454823Z digest=sha256:6eda84592b929f7a636b7c98fcc87f4b58050815f588a2548a4293a003f53178

Observation 6140211b-78a6-43e0-861f-4a3a7deab941 · inbound

Representation-Guided Parameter-Efficient LLM Unlearning cites this paper.

Representation-Guided Parameter-Efficient LLM Unlearning Knowledge Neurons in Pretrained Transformers

Reference 104

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:06:19.318927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T06:01:46.885030Z digest=sha256:2cdcc9fcd58fdf81098b7dfc288fb647b46f4962103c398327ce40dda35e3eed

Observation e552068b-6ea3-4b29-9aa9-8d4957107d3b · inbound

The Override Gap: A Magnitude Account of Knowledge Conflict Failure in Hypernetwork-Based Instant LLM Adaptation cites this paper.

The Override Gap: A Magnitude Account of Knowledge Conflict Failure in Hypernetwork-Based Instant LLM Adaptation Knowledge Neurons in Pretrained Transformers

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:11:10.247203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T06:39:00.768904Z digest=sha256:112e798d8795647356084f015d6aabd759755e63da62bcdd63d6bd0cc782ecdf

Observation f1e9da35-9720-463a-8b89-4401bc9c80b0 · inbound

The Override Gap: A Magnitude Account of Knowledge Conflict Failure in Hypernetwork-Based Instant LLM Adaptation cites this paper.

The Override Gap: A Magnitude Account of Knowledge Conflict Failure in Hypernetwork-Based Instant LLM Adaptation Knowledge Neurons in Pretrained Transformers

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:21:26.922333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T03:21:58.428036Z digest=sha256:72e24dda72b807fc7ab20ea37a1d80fc25ae0ce3f061482980efe933aa254079

Observation 291f1369-b205-4e69-8052-b59c779d33e0 · inbound

GKnow: Measuring the Entanglement of Gender Bias and Factual Gender cites this paper.

GKnow: Measuring the Entanglement of Gender Bias and Factual Gender Knowledge Neurons in Pretrained Transformers

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:52:16.737390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T04:50:43.547037Z digest=sha256:6fa43b5da71f174735fc25a1522214c0ee61b736e274e4cd103570ce02965b56

Observation b1c93311-09dd-467b-a3fc-f13222efeb15 · inbound

FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation cites this paper.

FocusDiT: Masking Queries in Diffusion Transformers for Fine-grained Image Generation Knowledge Neurons in Pretrained Transformers

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:26:17.391096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T15:25:40.039365Z digest=sha256:30dddc31145cf13098dec4a0608608c337d413b87b7ce390d2733bec63f3e988

Observation 166a21a5-bf99-41ed-8a3e-7a19b3991b3d · inbound

Why Muon Outperforms Adam: A Curvature Perspective cites this paper.

Why Muon Outperforms Adam: A Curvature Perspective Knowledge Neurons in Pretrained Transformers

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:06:44.945453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T07:04:21.012269Z digest=sha256:7e8505fa450babb9abe4b55967f92164aa69110d68297519c0ba157c03158bdd

Observation 28b5f54d-5508-44bf-9363-07afe0bb59e5 · inbound

Factual Retrieval in LLMs Is a Redundant, Distributed and Non-Contiguous Process cites this paper.

Factual Retrieval in LLMs Is a Redundant, Distributed and Non-Contiguous Process Knowledge Neurons in Pretrained Transformers

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:19:38.865535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T14:32:19.031688Z digest=sha256:8c51b8ba12f29c4bf1bf32e7a4032d12af1d542650288bc4233b70fe1038ce30

Observation 4da5e687-a263-4340-ba37-306d02f079e7 · inbound

Exposing the Illusion of Erasure in Knowledge Editing for LLMs cites this paper.

Exposing the Illusion of Erasure in Knowledge Editing for LLMs Knowledge Neurons in Pretrained Transformers

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:09:44.329523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T09:10:39.422141Z digest=sha256:543a4e531d1d0bb11c50ac8c1126d8a06c91a56182399cf9b6be4a804b81950a

Observation f790970d-bd63-4c59-a8ec-0d1b7f37df22 · inbound

Seeing Through Multiple Views: Parameter-Efficient Fine-Tuning via Selective Neurons for Consistent Radiology Report Generation cites this paper.

Seeing Through Multiple Views: Parameter-Efficient Fine-Tuning via Selective Neurons for Consistent Radiology Report Generation Knowledge Neurons in Pretrained Transformers

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:45:40.552266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T06:10:32.156961Z digest=sha256:c9297fef24468f1503ac391965362bd7fa131f82ccb686d826a891bc389b3f74

Observation cdf0d6e8-40f9-460f-8e5a-8d27ecab3064 · inbound

Break Through the Compression Bottleneck: From Theory to Practice cites this paper.

Break Through the Compression Bottleneck: From Theory to Practice Knowledge Neurons in Pretrained Transformers

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T14:29:31.857953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T14:29:31.857953Z digest=sha256:c8270d435d7c3f062a7a8151d123f2c6fd8045bb93d05cf999d570eac49c33c0

Observation 8095d1a4-8f2e-405c-b549-a5491e2c186d · inbound

Context-Adaptive Inference: A Unified Statistical and Foundation-Model View cites this paper.

Context-Adaptive Inference: A Unified Statistical and Foundation-Model View Knowledge Neurons in Pretrained Transformers

Reference 204

Resolution
unresolved
no resolver link, observed 2026-07-31T23:53:01.327684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:53:01.327684Z digest=sha256:19a8d8434cb0ad6e93801491cc25e673b1aeff23200dc400e5aa0f7d5ef874d0

Observation 49ef6bcf-eec6-47e2-9a44-36c33787a5d0 · inbound

Fairness Pruning: Locating Demographic Bias in GLU-MLP Layers via Differential Activations cites this paper.

Fairness Pruning: Locating Demographic Bias in GLU-MLP Layers via Differential Activations Knowledge Neurons in Pretrained Transformers

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-31T11:25:39.503891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T11:25:39.503891Z digest=sha256:05c91cfcf833a1a13690329b7f6c843d1694c5259671f973463ae2c4eade87bf