Pith. sign in

Paper Citation Record · LEDGER

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders

As of 12 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 0 inbound Pith citation observations for arXiv:2608.08168.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.08168 v1

Coverage vector

measured 27 of 27 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:23:50.523918Z

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

27 of 27 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9cbcc900-a783-4fb0-9e27-92102fc38583 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.418419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.418419Z digest=sha256:bb45fe983983501c1fa91a77c17dc39ede885db546abf58a438bce8f64875076

Observation 18ea0f4e-bc63-4f89-859c-f3711b44f138 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.423415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.423415Z digest=sha256:39d05079093b07a53e5d84250d35f7944b90f00e79aec112ca8b1e95756c0305

Observation 1b8f143f-e2be-4612-b50f-c6a9e03c2836 · outbound

This paper cites Language models don’t always say what they think: Unfaithful explanations in chain-of-thought prompting.Advances in Neural Information Processing Systems, 36:74952–74965, 2023.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Language models don’t always say what they think: Unfaithful explanations in chain-of-thought prompting.Advances in Neural Information Processing Systems, 36:74952–74965, 2023

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.427827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.427827Z digest=sha256:673e1c76e670fae53eaed7dbe8aa0b43cb988456222a0522ec1da175990404e1

Observation 290a74bd-ce7a-4316-875e-ad22de85c5d3 · outbound

This paper cites Let’s verify step by step.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Let’s verify step by step

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.431998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.431998Z digest=sha256:452cd13c211fc49e8f5453bdff4926d49dee9d7078cd0d5fb2df01d8a169513d

Observation 33f4f543-12ee-479c-82d3-1435710c04de · outbound

This paper cites How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders How does Chain of Thought Think? Mechanistic Interpretability of Chain-of-Thought Reasoning with Sparse Autoencoding

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.436202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.436202Z digest=sha256:9395a4b5be4f5a47ca1f3f01dd781cdab7195c9a50b81b4be5cc877f33ac2424

Observation 25f08bb4-d134-458b-8514-0b0de5a37b2b · outbound

This paper cites Finding sparse autoencoder representations of errors in cot prompting.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Finding sparse autoencoder representations of errors in cot prompting

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:23:50.891768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T00:23:50.441305Z digest=sha256:624a88280d29f654fa794cb957f6388617515e87f53ef1fd9576beaf1ac35676

Observation 2d4e165a-6dc8-4c43-90b0-a2c62a911c0e · outbound

This paper cites Progress measures for grokking via mechanistic interpretability.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Progress measures for grokking via mechanistic interpretability

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.445426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.445426Z digest=sha256:f150ab254a512292b455c1f0b178ecce2334e5a6c46a3ac0bff9c9fbf4c09788

Observation 3007f258-dc91-4983-88cd-871ce36b1c59 · outbound

This paper cites Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.449277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.449277Z digest=sha256:eb57beeb6edf67700aec1760428cd7f94b46d9549fb35c59c15b9b0055d2c0f8

Observation 87bd8101-a35a-44b4-92c1-0f84e1adcae6 · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Representation Engineering: A Top-Down Approach to AI Transparency

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.453594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.453594Z digest=sha256:20ebb7a06c0f75a6e30d9fd13d90a91cd19d18e7af976a415d55551131a43e95

Observation 5a1bb2e0-1002-4714-95ed-e3b8717a381a · outbound

This paper cites Improving Dictionary Learning with Gated Sparse Autoencoders.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Improving Dictionary Learning with Gated Sparse Autoencoders

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.457846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.457846Z digest=sha256:2b7fe810647a886ba1c6da037ec04ac4f02799a27447f6a56bb6e9b39cbbd1e7

Observation 8328cf95-7b15-438b-b4bb-a9189be04b13 · outbound

This paper cites Sparse Autoencoders Find Highly Interpretable Features in Language Models.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Sparse Autoencoders Find Highly Interpretable Features in Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.462347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.462347Z digest=sha256:3b9359d416ba09e7517ddb4b4dca21859dcf33ec880eda3aa2e684fe8d8fc0a9

Observation dc9a124f-eb76-40b0-83f5-827bb157a498 · outbound

This paper cites Towards monosemanticity: Decomposing language models with dictionary learning.Transformer Circuits Thread, 2, 2023.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Towards monosemanticity: Decomposing language models with dictionary learning.Transformer Circuits Thread, 2, 2023

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.466491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.466491Z digest=sha256:9a3546ee00126a62e95f32e934818d4a15f7922e6bff97a7a2bd063d888f3ae4

Observation 115a9d7f-3f22-4362-9688-271f013aea15 · outbound

This paper cites Feature Extraction and Steering for Enhanced Chain-of-Thought Reasoning in Language Models.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Feature Extraction and Steering for Enhanced Chain-of-Thought Reasoning in Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.470682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.470682Z digest=sha256:5a27a497408a3dcac4244ed734fa257e2b36f00c78b97ecdd90841ecaeb9ada9

Observation 6b0e1e58-301d-4bf1-ae33-d7a88e444086 · outbound

This paper cites Locating and editing factual associations in gpt.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Locating and editing factual associations in gpt

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.475049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.475049Z digest=sha256:e8d78322d70695672e542bc2d98b8aa182b5ba6a145c0fbcadbd59610a9660fa

Observation 27b113c0-62ab-4821-9d15-3119b15046de · outbound

This paper cites Scaling and evaluating sparse autoencoders.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Scaling and evaluating sparse autoencoders

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.478491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.478491Z digest=sha256:f431a72579ebba2986c758d07929006e05f9159777b5820ac997e5679e38d634

Observation ada82d09-0663-421c-a985-c39ccd67f7e3 · outbound

This paper cites Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.482648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.482648Z digest=sha256:bb79dbe9aaa62865274e22793548574561621eb9dbc1fa8bfc57111f390e4b65

Observation 62dbf1e2-29f5-4e4e-8a34-db0cdfac515b · outbound

This paper cites Sparse autoencoder features for classifications and transferability.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Sparse autoencoder features for classifications and transferability

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:23:50.864087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T00:23:50.486695Z digest=sha256:6c271022e9b5ac76e4eb201c9256bfdfbfef43d09aa6bab142bb3c38b6354f65

Observation 27f7208c-d0eb-4fdf-806f-3f14ac1c762e · outbound

This paper cites Saes are good for steering–if you select the right features.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Saes are good for steering–if you select the right features

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.490578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.490578Z digest=sha256:5bbc5c86a86e9362b6cacc072a471a341c86234e7959e0b7a0bd973f39983aef

Observation ece60cbd-713e-4dd9-a3dd-e999016a653b · outbound

This paper cites Lingualens: Towards interpreting linguistic mechanisms of large language models via sparse auto-encoder.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Lingualens: Towards interpreting linguistic mechanisms of large language models via sparse auto-encoder

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:23:50.851200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T00:23:50.494278Z digest=sha256:3ef055dc21053e38d460ebf90ea3c45c80ebf86730e631a2f19b993b898c641a

Observation 14b9792e-6e85-4ffc-87fe-a920bfaaa251 · outbound

This paper cites Decoding Dense Embeddings: Sparse Autoencoders for Interpreting and Discretizing Dense Retrieval.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Decoding Dense Embeddings: Sparse Autoencoders for Interpreting and Discretizing Dense Retrieval

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.498053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.498053Z digest=sha256:7d6512612c448f1b943257aa5f199f4319b93caeffa4a97d9c00b679b949cb10

Observation 9a9da866-17ca-4060-9735-6573655149a4 · outbound

This paper cites k-Sparse Autoencoders.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders k-Sparse Autoencoders

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.501902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.501902Z digest=sha256:93e2aad19e52e4c960b73e775c1dc7a40652dbd64ccf790c2949ccc4643f9448

Observation 263e575a-3be9-4270-84bf-193f546a2271 · outbound

This paper cites DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.505227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.505227Z digest=sha256:3fdde5d23889334f3bd99ab262af78328c3f89e4702b5a836d9f19536246cb1f

Observation e503d8e4-528e-487f-9c43-5f1f235040d1 · outbound

This paper cites Reasoning Models Can Be Effective Without Thinking.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Reasoning Models Can Be Effective Without Thinking

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.508557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.508557Z digest=sha256:2f37beac665dc9a6b323a22f66321f5d9c4a79c46650cd660e8140b2e6e8b2e0

Observation 53ef2980-cb57-4d52-83ca-8bf913379abd · outbound

This paper cites Interpreting and steering llm representations with mutual information-based explanations on sparse autoencoders.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Interpreting and steering llm representations with mutual information-based explanations on sparse autoencoders

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:23:50.836653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T00:23:50.512555Z digest=sha256:2bedb24b94be9a0d9245731c48f5873c9874ad17e68bbfbe021578511e2e404d

Observation b85240b3-6c73-4274-b747-c8c9a8e3cf3d · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Adam: A Method for Stochastic Optimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T00:23:50.516505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:23:50.516505Z digest=sha256:5d8296c20028170634f3a19ded298277c22c6b775a86fa762d0d0fd7e67c6fad

Observation eb2de8b2-28f3-4360-a468-ce05979d74fb · outbound

This paper cites Scaling monosemanticity: Extracting interpretable features from claude 3 sonnet.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Scaling monosemanticity: Extracting interpretable features from claude 3 sonnet

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:23:50.821013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T00:23:50.520191Z digest=sha256:7cd353ee21b1378c36ccfb02ff58737bbfa300eb2e8a371abf10e53834b3a8ca

Observation a8016092-872c-4628-bf30-8c059e3c4502 · outbound

This paper cites wait", "hmm.

Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders wait", "hmm

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T00:23:50.806064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T00:23:50.523918Z digest=sha256:1ad0b0a80ae4dbc6d2a1670ce147a9901e2dc3b388e5dbb7584c4310c57f5397

Pith citing papers

No inbound Pith citation observations are available.