Pith. sign in

Paper Citation Record · LEDGER

A Primer on the Inner Workings of Transformer-based Language Models

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 47 inbound Pith citation observations for arXiv:2405.00208.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.00208 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 47 of 47 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:19:44.196854Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

9
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 926a3091-f99a-4de3-816c-decb566fdf78 · inbound

Do I Know This Entity? Knowledge Awareness and Hallucinations in Language Models cites this paper.

Do I Know This Entity? Knowledge Awareness and Hallucinations in Language Models A Primer on the Inner Workings of Transformer-based Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T15:28:21.437981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T15:28:21.437981Z digest=sha256:caadf3a2f0832af799b914f2416b877e9358b5d6080890a6930f5b3d6247b202

Observation 2b038d02-bc95-459d-9304-50022b173da5 · inbound

Evaluating Sparse Autoencoders on Targeted Concept Erasure Tasks cites this paper.

Evaluating Sparse Autoencoders on Targeted Concept Erasure Tasks A Primer on the Inner Workings of Transformer-based Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T10:51:14.246042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:51:14.246042Z digest=sha256:b741339e621bf7c7e19bc19681887e4f3a66c4914b7f61d92c1bdaeb450fd669

Observation 88c67020-6368-4df0-95b5-e3a433e48f2c · inbound

RevPRAG: Revealing Poisoning Attacks in Retrieval-Augmented Generation through LLM Activation Analysis cites this paper.

RevPRAG: Revealing Poisoning Attacks in Retrieval-Augmented Generation through LLM Activation Analysis A Primer on the Inner Workings of Transformer-based Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T10:47:45.713695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T10:47:45.713695Z digest=sha256:d70439ec5d898daf30c340ef868e5cc5d1c4c007ac5a0b2274d2d78a82c803df

Observation 0ca94909-1577-4538-b09f-ab36cb84e724 · inbound

Inferring Functionality of Attention Heads from their Parameters cites this paper.

Inferring Functionality of Attention Heads from their Parameters A Primer on the Inner Workings of Transformer-based Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T14:29:42.761225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:29:42.761225Z digest=sha256:acea85b3eab2f251515904df28478e8015c7c6a678360c6de21e06a19f462074

Observation 1052c760-23d3-4661-b152-2d00b615b01d · inbound

Towards scientific discovery with dictionary learning: Extracting biological concepts from microscopy foundation models cites this paper.

Towards scientific discovery with dictionary learning: Extracting biological concepts from microscopy foundation models A Primer on the Inner Workings of Transformer-based Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-11T11:29:14.896147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:29:14.896147Z digest=sha256:3f9786bbbd92f9f36d07fd124317ec87b83e22e78f4932cb258b253c6adb71d7

Observation 3724ba54-01c3-42c5-9171-8a6060021378 · inbound

Correctness Assessment of Code Generated by Large Language Models Using Internal Representations cites this paper.

Correctness Assessment of Code Generated by Large Language Models Using Internal Representations A Primer on the Inner Workings of Transformer-based Language Models

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-10T16:41:31.340644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:41:31.340644Z digest=sha256:c6902d82a9f55c6071813b177c9ae29d1b5fb6169abfd19e2ed0c2751636efda

Observation 4c290d7f-f079-48b8-848c-2dcec4755adb · inbound

SEAL: Scaling to Emphasize Attention for Long-Context Retrieval cites this paper.

SEAL: Scaling to Emphasize Attention for Long-Context Retrieval A Primer on the Inner Workings of Transformer-based Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T14:32:50.043146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T14:32:50.043146Z digest=sha256:565357c292e045b1e95af5afbb40990fe5ed5d1eee9c6fd442d867b4efaa84bb

Observation 2cb92199-7342-4f6f-b00b-af85497cc4c4 · inbound

MIB: A Mechanistic Interpretability Benchmark cites this paper.

MIB: A Mechanistic Interpretability Benchmark A Primer on the Inner Workings of Transformer-based Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T12:19:44.196854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T12:19:44.196854Z digest=sha256:9bb64a3abbdad0a2dd53142dd454b183eb56f5f2dd457996ceabc6b81dce8296

Observation 07f6ca5a-e9e3-4625-a50a-1fc7a6ec5729 · inbound

Lightweight Latent Verifiers for Efficient Meta-Generation Strategies cites this paper.

Lightweight Latent Verifiers for Efficient Meta-Generation Strategies A Primer on the Inner Workings of Transformer-based Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T11:00:20.377178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:00:20.377178Z digest=sha256:ace6b8dd94906161627dd3fa073cfa098c49696b13634ef05c3899f5e57771e4

Observation e0053fa9-4672-48a9-8a63-c0052c2ea43e · inbound

Detection and Mitigation of Hallucination in Large Reasoning Models: A Mechanistic Perspective cites this paper.

Detection and Mitigation of Hallucination in Large Reasoning Models: A Mechanistic Perspective A Primer on the Inner Workings of Transformer-based Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T20:28:18.967260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:28:18.967260Z digest=sha256:100f82a4262c7000fd662f17fb5a45e31e2f05e07477bd7c192b0488d2f0557f

Observation cd891a5e-fdcc-4edc-aee8-29bd0f28be5b · inbound

Localizing Persona Representations in LLMs cites this paper.

Localizing Persona Representations in LLMs A Primer on the Inner Workings of Transformer-based Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:25:31.903456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:25:31.903456Z digest=sha256:f4a79923ac222a91559c8b03b058735c614c11add51de4f25ce1fb670276103f

Observation ff9321fb-f648-45d5-911f-03a2f7b310df · inbound

COMPKE: Complex Question Answering under Knowledge Editing cites this paper.

COMPKE: Complex Question Answering under Knowledge Editing A Primer on the Inner Workings of Transformer-based Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:20.644602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:00:20.644602Z digest=sha256:ec047a88828608eccba8352ec4f7d59d07376f6aea431eaab380bebda4c30565

Observation 5d965c15-1fd4-4e84-95fd-5844c1471ab9 · inbound

Different Speech Translation Models Encode and Translate Speaker Gender Differently cites this paper.

Different Speech Translation Models Encode and Translate Speaker Gender Differently A Primer on the Inner Workings of Transformer-based Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:41.614831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:32:41.614831Z digest=sha256:18df6bae38f4600c7d469e6424383d24efa48dfbdca9bf0b5ce37223614c408f

Observation a7746e88-e267-4a16-ab21-a2c6f35cccb0 · inbound

REAL: Reading Out Transformer Activations for Precise Localization in Language Model Steering cites this paper.

REAL: Reading Out Transformer Activations for Precise Localization in Language Model Steering A Primer on the Inner Workings of Transformer-based Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:30.100608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:19:30.100608Z digest=sha256:5c4f40fab6bcc136fd3658b063cabfb00ce2d437e1b14e3b94524e4e2fd2eb12

Observation d039dfbb-1351-4c27-869b-4fa661590064 · inbound

Private Memorization Editing: Turning Memorization into a Defense to Strengthen Data Privacy in Large Language Models cites this paper.

Private Memorization Editing: Turning Memorization into a Defense to Strengthen Data Privacy in Large Language Models A Primer on the Inner Workings of Transformer-based Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:08.613581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:27:08.613581Z digest=sha256:c8ee439f272373e96658a94f8fc7442f0c8c275e0e586b79f3fef002b9af39e7

Observation 026616c2-d0a3-4157-b512-33a7d272602d · inbound

A Free Probabilistic Framework for Analyzing the Transformer-based Language Models cites this paper.

A Free Probabilistic Framework for Analyzing the Transformer-based Language Models A Primer on the Inner Workings of Transformer-based Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T19:31:29.473545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:31:29.473545Z digest=sha256:f2cba6df5b9cc02a13baab29ca8be3026fd0e730b258d74c0c79c81b9c124d57

Observation 969dbf6a-d78c-4991-b7cc-d00e476c2663 · inbound

Table Understanding and (Multimodal) LLMs: A Cross-Domain Case Study on Scientific vs. Non-Scientific Data cites this paper.

Table Understanding and (Multimodal) LLMs: A Cross-Domain Case Study on Scientific vs. Non-Scientific Data A Primer on the Inner Workings of Transformer-based Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:26:12.059928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:26:12.059928Z digest=sha256:2d7954b2f960e19798a035ed52f0f59ecb8579c4817b336b54c501296dc4aa28

Observation a7107925-66e7-4d5e-939b-6ac8c480cc71 · inbound

Reconstructing Biological Pathways by Applying Selective Incremental Learning to (Very) Small Language Models cites this paper.

Reconstructing Biological Pathways by Applying Selective Incremental Learning to (Very) Small Language Models A Primer on the Inner Workings of Transformer-based Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T19:53:37.628279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:53:37.628279Z digest=sha256:d4281b8e1233d3b4b50f13f7a408c8714e1bda735f8adfa6f9b2c63c6776c28a

Observation 33e50dca-6d09-4305-be81-0a65a7de62a0 · inbound

InTraVisTo: Inside Transformer Visualisation Tool cites this paper.

InTraVisTo: Inside Transformer Visualisation Tool A Primer on the Inner Workings of Transformer-based Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T16:19:52.229487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:19:52.229487Z digest=sha256:032155e4431f247be415c3fadcf11148cc6c160b0b22f1508d51da078fcceab1

Observation 6360b106-6c69-4543-bebd-268732f5067c · inbound

Cross-Attention is Half Explanation in Speech-to-Text Models cites this paper.

Cross-Attention is Half Explanation in Speech-to-Text Models A Primer on the Inner Workings of Transformer-based Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T15:52:06.284744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T15:52:06.284744Z digest=sha256:fab59b6c796209b45a498fe41d83c95a251cef6a8f0a4ecb8ff49019205bc3ac

Observation ae9637fe-1be0-4efa-9434-4a352bf193e0 · inbound

Safe-SAIL: Towards a Fine-grained Safety Landscape of Large Language Models via Sparse Autoencoder Interpretation Framework cites this paper.

Safe-SAIL: Towards a Fine-grained Safety Landscape of Large Language Models via Sparse Autoencoder Interpretation Framework A Primer on the Inner Workings of Transformer-based Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T18:16:43.774882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-18T18:13:01.662828Z digest=sha256:ee10616e6feb55c8aa6e50310ae6e548a829b5b069a30ab1d0faf509175cedd5

Observation 4dd6da5a-0fac-4782-889b-93e2cb759d3b · inbound

Deductive Logic in Language Models: Horizontal vs Vertical Reasoning cites this paper.

Deductive Logic in Language Models: Horizontal vs Vertical Reasoning A Primer on the Inner Workings of Transformer-based Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T10:40:26.442823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:40:26.442823Z digest=sha256:6e4df0be96762ba9ccaa8ed0cbd5033b65be3a1ccccb600d96ece6dde531741b

Observation f1af8c93-d940-408c-9025-3865fb76a609 · inbound

Explainability of Large Language Models: Opportunities and Challenges toward Generating Trustworthy Explanations cites this paper.

Explainability of Large Language Models: Opportunities and Challenges toward Generating Trustworthy Explanations A Primer on the Inner Workings of Transformer-based Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:54.428626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:54.428626Z digest=sha256:8f68cc9075548e15997e1c74340ed1c16d7323f2fcaf75f333339040dde12049

Observation 11c8130f-c5be-4e8d-8cab-00bd528aee42 · inbound

Unifying Learning Dynamics and Generalization in Transformers Scaling Law cites this paper.

Unifying Learning Dynamics and Generalization in Transformers Scaling Law A Primer on the Inner Workings of Transformer-based Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T14:02:56.666237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:02:56.666237Z digest=sha256:b1ae5b25e7d70ef05d13d7c1285c80a1ce89a7b28c63367ab86dc98f84b804e8

Observation f53c37eb-ce5e-4382-a38d-c96bc8906a9a · inbound

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models cites this paper.

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models A Primer on the Inner Workings of Transformer-based Language Models

Reference 72

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:40:54.748973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T12:39:57.398423Z digest=sha256:034bbcd26bab98fdea3900a87a6d56da49a2240bc64a4cf8c74abcfc1659befe

Observation 09abdd37-4ecf-4c6e-83e9-0eb4e5e2abb6 · inbound

MICE: Minimal Interaction Cross-Encoders for efficient Re-ranking cites this paper.

MICE: Minimal Interaction Cross-Encoders for efficient Re-ranking A Primer on the Inner Workings of Transformer-based Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-02T22:39:09.674583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:39:09.674583Z digest=sha256:4525caf66bed70e5e2e110ffdf1bc61c05df659b3d20f4514c2734e7038f8d6d

Observation 026ccf86-4225-4276-8128-7c829353877f · inbound

Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models cites this paper.

Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models A Primer on the Inner Workings of Transformer-based Language Models

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-21T12:50:09.275117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-21T12:46:59.419162Z digest=sha256:82f5887e2620c94999e4cee78ea2bb20b346f3fa1a1a3febb5af82ce4748baab

Observation 10668668-0420-40c4-98e8-0eed8aa1a1cf · inbound

In-Context Learning in Speech Language Models: Analyzing the Role of Acoustic Features, Linguistic Structure, and Induction Heads cites this paper.

In-Context Learning in Speech Language Models: Analyzing the Role of Acoustic Features, Linguistic Structure, and Induction Heads A Primer on the Inner Workings of Transformer-based Language Models

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:05:49.019611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T20:14:21.150863Z digest=sha256:5950654ef71cf7ab49e1a518434f74658da89ae789d95ba657e7f66809482a56

Observation 83663496-d765-4e1c-a429-25342bb7063c · inbound

ConceptTracer: Interactive Analysis of Concept Saliency and Selectivity in Neural Representations cites this paper.

ConceptTracer: Interactive Analysis of Concept Saliency and Selectivity in Neural Representations A Primer on the Inner Workings of Transformer-based Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:20:51.372895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T19:13:21.974912Z digest=sha256:a5bd3db60b4cd9fd7bf01957c66b8bccdc4133214ce112bddc5d48638220b188

Observation 25a79f0f-dd89-4954-bebd-aff1becd1fab · inbound

Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks cites this paper.

Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks A Primer on the Inner Workings of Transformer-based Language Models

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:38:42.816884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T05:12:31.218055Z digest=sha256:2156a19f75d96251b3c79edbd2bd0af62095b6c88c3e124ef52b11cfd794228f

Observation c562ffef-6152-4ae4-b75a-74d28d7c9c9d · inbound

Are LLMs Ready for Conflict Monitoring? Empirical Evidence from West Africa cites this paper.

Are LLMs Ready for Conflict Monitoring? Empirical Evidence from West Africa A Primer on the Inner Workings of Transformer-based Language Models

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:11:18.021348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-08T17:48:51.792769Z digest=sha256:7102707ab42370a63d7436ce8753ddc896376e0293c17c1d12649aa9b4749fc9

Observation 6cc7a763-e24b-4086-b8fd-81fcb49e69ec · inbound

Data-driven Circuit Discovery for Interpretability of Language Models cites this paper.

Data-driven Circuit Discovery for Interpretability of Language Models A Primer on the Inner Workings of Transformer-based Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:21:16.662417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T02:19:45.010301Z digest=sha256:4feae379340e3e80267bce71b19aa8a9d5068dc1863e781c25af51c4c8309d1c

Observation 3ee19c8b-1c38-4afc-a82f-f9003397b20b · inbound

Enabling Performant and Flexible Model-Internal Observability for LLM Inference cites this paper.

Enabling Performant and Flexible Model-Internal Observability for LLM Inference A Primer on the Inner Workings of Transformer-based Language Models

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:17:28.083335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T07:17:13.823853Z digest=sha256:f0c48a9492d343dc142496328285533ef39d1a9e47c8b1bc1cb8697f148d2b30

Observation f714c5e2-0d39-4939-a70c-b660c5e12355 · inbound

Instructions Shape Production of Language, not Processing cites this paper.

Instructions Shape Production of Language, not Processing A Primer on the Inner Workings of Transformer-based Language Models

Reference 138

Resolution
verified exact
arxiv_id, observed 2026-05-13T03:12:09.296748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-13T03:09:02.902912Z digest=sha256:6024d5e1efc9ef7ef28ce7ea44cdbc9c863ba90b3d30c3211b756b5b1e357089

Observation d957face-3a0b-471f-9aa8-effa8058e3e4 · inbound

Instructions Shape Production of Language, not Processing cites this paper.

Instructions Shape Production of Language, not Processing A Primer on the Inner Workings of Transformer-based Language Models

Reference 138

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:02:58.204816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-14T21:02:02.135970Z digest=sha256:84575be479ef67ffe99492b8ae0132fe1e5db306e4d1b9c4c2f12bffe6e796f3

Observation d20cb649-d16e-429e-8140-99aecf524e17 · inbound

Stories in Space: In-Context Learning Trajectories in Conceptual Belief Space cites this paper.

Stories in Space: In-Context Learning Trajectories in Conceptual Belief Space A Primer on the Inner Workings of Transformer-based Language Models

Reference 74

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:22:18.929668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-13T05:17:34.283917Z digest=sha256:c816bbf088637afc4c8f5049770003d8536467fbe14dd28e3dbf5d433a0638d9

Observation 676e7e43-ddd8-49e9-8fcd-52512a69dd4c · inbound

Rethinking Layer Relevance in Large Language Models Beyond Cosine Similarity cites this paper.

Rethinking Layer Relevance in Large Language Models Beyond Cosine Similarity A Primer on the Inner Workings of Transformer-based Language Models

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:09:44.376483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-15T05:05:01.099084Z digest=sha256:1bc03ca3bb02296a7d7e71788b9eef679fcc95df4cfc8651b66eaafab36f5f76

Observation 6e249306-266e-478a-ac6b-c035aa77ec84 · inbound

Relational Linear Properties in Language Models: An Empirical Investigation cites this paper.

Relational Linear Properties in Language Models: An Empirical Investigation A Primer on the Inner Workings of Transformer-based Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:46:14.986084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-22T07:45:09.743615Z digest=sha256:55053d38e398b2e434f8bc6a47b4b5eec3b7262c77353cd55c8d7e4a98036041

Observation d9ed1830-ba6f-45ca-8190-44260ed1cec7 · inbound

Relational Linear Properties in Language Models: An Empirical Investigation cites this paper.

Relational Linear Properties in Language Models: An Empirical Investigation A Primer on the Inner Workings of Transformer-based Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T17:14:56.759886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T17:13:03.620676Z digest=sha256:9897053acf518220f1914883c8b687bb0b250ef22ef18fa3c535a0a56919ad7d

Observation 98d98515-5b4a-422d-aa76-daff08a04ab8 · inbound

Multi-component Causal Tracing in Large Language Models cites this paper.

Multi-component Causal Tracing in Large Language Models A Primer on the Inner Workings of Transformer-based Language Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:36:24.919128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T11:52:24.954950Z digest=sha256:495e746f9fbf9640dd6b75dec7eca1e4635f352039217fa79563814d340c64ee

Observation 8a866893-c40e-4719-84fb-e708dbbc3f9e · inbound

Contribution Weights: A Geometrical Analysis of Self-Attention Transformers cites this paper.

Contribution Weights: A Geometrical Analysis of Self-Attention Transformers A Primer on the Inner Workings of Transformer-based Language Models

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:32:46.752725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T23:29:02.457697Z digest=sha256:dbe0751810e0fe0dfa7e6467e2fa9c3f3ce1cec89c127d7a218e0412d5a60ea9

Observation aa70353b-9fe2-4354-843f-4502b7615ab5 · inbound

Contribution Weights: A Geometrical Analysis of Self-Attention Transformers cites this paper.

Contribution Weights: A Geometrical Analysis of Self-Attention Transformers A Primer on the Inner Workings of Transformer-based Language Models

Reference 158

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:32:47.304588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T23:29:02.457697Z digest=sha256:6b31acd205a7e91cf14569912389d998705cbc173dedaa4b18145c6d98d9b4cf

Observation 17b1d4fc-ca74-41a8-81d9-4e6feb7269e5 · inbound

DECSELFMASK: Leveraging Unlabeled Text via Self-Relevance-Guided Masking for Decoder-Only Classification cites this paper.

DECSELFMASK: Leveraging Unlabeled Text via Self-Relevance-Guided Masking for Decoder-Only Classification A Primer on the Inner Workings of Transformer-based Language Models

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T01:27:30.553886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T16:36:10.784902Z digest=sha256:1472290c7d428402e0c4bca800e8fb5d78e20b4756b12da0d6a4411b9d66c894

Observation 630106ce-c732-46c8-a961-19f73307ba2a · inbound

Can Language Model Agents be Helpful Circuit Explainers in Mechanistic Interpretability? cites this paper.

Can Language Model Agents be Helpful Circuit Explainers in Mechanistic Interpretability? A Primer on the Inner Workings of Transformer-based Language Models

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T16:19:57.641510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T00:47:00.310205Z digest=sha256:91989cfef40ba66d957c7571cd7dad158a6114848b1c44c9b5b72e1c35296e9f

Observation 6489dc32-5b72-4f04-a3da-eb0acccce6a4 · inbound

PRA-RAG: Provably Robust Aggregation in Retrieval-Augmented Generation against Retrieval Corruption cites this paper.

PRA-RAG: Provably Robust Aggregation in Retrieval-Augmented Generation against Retrieval Corruption A Primer on the Inner Workings of Transformer-based Language Models

Reference 65

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T23:47:27.188119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-07-02T23:42:23.695930Z digest=sha256:66cef71597e5496ac234437968a9abae1bbf27d93a6395c52d6f4b469a13f6e6

Observation b8170ba8-a3cf-4660-ae68-63ce9d7b66b4 · inbound

Weight-Adjusted Gradients Reveal Parameter Importance and Failure Modes in LLMs cites this paper.

Weight-Adjusted Gradients Reveal Parameter Importance and Failure Modes in LLMs A Primer on the Inner Workings of Transformer-based Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-14T09:09:29.470093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T09:09:29.470093Z digest=sha256:87b410c82e4a3bd7d883471d9cfbb3f64281739973ebe752d65bc9709df1228d

Observation 497a975c-171a-4512-a7e8-a9c3b82eb8c6 · inbound

On the use of foundation models in cognitive science cites this paper.

On the use of foundation models in cognitive science A Primer on the Inner Workings of Transformer-based Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T04:16:34.075212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:16:34.075212Z digest=sha256:4ae85cde5f79597af9f30c9e98948f4299913dbc26c8010320fe2bcf515d7585