Pith. sign in

Paper Citation Record · LEDGER

How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained language model

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2305.00586.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.00586 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:13:17.160012Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

14
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 977a0f4c-402a-4e83-b388-7ed56044895c · inbound

How to use and interpret activation patching cites this paper.

How to use and interpret activation patching How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained language model

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:34:06.964524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T20:34:06.888683Z digest=sha256:988a3d7ade0274752b64aabde15554e23bf0ff7ea89a6c85998712ba070b55e9

Observation b85fd806-9985-45e6-81e9-159e2ab2f9e0 · inbound

Aligned but Blind: Alignment Increases Implicit Bias by Reducing Awareness of Race cites this paper.

Aligned but Blind: Alignment Increases Implicit Bias by Reducing Awareness of Race How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained language model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:17.160012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:13:17.160012Z digest=sha256:6703c4c4eaaba078f914c361e39df65a15341f017f936f23e470962f76614e29

Observation 0886e2a4-bb5e-4221-87ee-b77f5ad9c369 · inbound

NEAT: Concept driven Neuron Attribution in LLMs cites this paper.

NEAT: Concept driven Neuron Attribution in LLMs How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained language model

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T18:01:19.631659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:01:19.631659Z digest=sha256:9bd71b908f046128540861bf443fb9276d3fb7946a96a943f18d0df5fa04f885

Observation a5296ca8-d6ab-4752-b351-83d9a87d7f96 · inbound

From Indirect Object Identification to Syllogisms: Exploring Binary Mechanisms in Transformer Circuits cites this paper.

From Indirect Object Identification to Syllogisms: Exploring Binary Mechanisms in Transformer Circuits How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained language model

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T17:37:53.812763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:37:53.812763Z digest=sha256:ad9ae2b4df2e766a8790d44cf24f65ca7260824b0d5a124ee1829aa0927cf26a

Observation b9c19a30-250e-411e-824e-e509d3b75155 · inbound

Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning cites this paper.

Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained language model

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T13:03:22.096173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:03:22.096173Z digest=sha256:1a6c28c63d985f81d32bc4f03ac4b043f327bc4452b38a5e7ef40726f1d23cd9

Observation cc775181-050f-48b9-953f-0384c27e84f6 · inbound

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces cites this paper.

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained language model

Reference 152

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:17:54.449139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-14T20:17:01.224864Z digest=sha256:b896bb9cf49a790a35b5055a0150bd22750d0d266efc9a58154d26fc51fee0c4

Observation 60b0f94f-d1f6-4d2e-be43-50e1a9f8b02d · inbound

Pattern Selectivity is Not Task-Causal Structure: A Cross-Architecture Mechanistic Study of Composed-Task Circuits in 1B-Class Language Models cites this paper.

Pattern Selectivity is Not Task-Causal Structure: A Cross-Architecture Mechanistic Study of Composed-Task Circuits in 1B-Class Language Models How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained language model

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:06:44.475064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T07:07:58.170765Z digest=sha256:692217400cbd62681c6bf667972ffcd12160195f757ae0d7302a60349bb630a8

Observation 605507d4-3db6-4515-9abd-a7df0d789ecf · inbound

Can Language Model Agents be Helpful Circuit Explainers in Mechanistic Interpretability? cites this paper.

Can Language Model Agents be Helpful Circuit Explainers in Mechanistic Interpretability? How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained language model

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T16:19:57.638846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T00:47:00.310205Z digest=sha256:c2760c207cdc550799a61499a7789da1c20e5db971d56d396466c045ff7dc27c

Observation 95278f95-773d-4c71-a00a-1c1eaf1caf37 · inbound

Pretraining Curricula Enable Selective Fine-tuning cites this paper.

Pretraining Curricula Enable Selective Fine-tuning How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained language model

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-11T12:36:24.747752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T12:36:24.747752Z digest=sha256:2846cb3c7828dad857c64c8c91d68d17b63d38352a4d669b1559b437e49307df