Pith. sign in

Paper Citation Record · LEDGER

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs

As of 4 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 1 inbound Pith citation observation for arXiv:2602.10352.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.10352 v2

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T01:14:59.660318Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-12T03:21:30.932283Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T07:26:26.034345Z

Reference resolution

20 of 20 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved18
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ed057f1e-cdae-4fdf-bc73-325d66801d0c · outbound

This paper cites Tell me about.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Tell me about

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:58.647413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:58.647413Z digest=sha256:7feb355ef20726e434de087131ef450a4436dad2b2f2dd1b2dbe80f748729d5c

Observation 1b8e4095-0c9a-4f18-875c-7f11a7526bd9 · outbound

This paper cites Factorials in combinatorics.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Factorials in combinatorics

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:58.718966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:58.718966Z digest=sha256:01cf089fc3f7dd7716ee6597f8f3f5600322381dfa09597892a8ade97d8f75f7

Observation 4c2757a0-0aa2-470a-bc9b-89cf7b502842 · outbound

This paper cites Tell me about [topic].

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Tell me about [topic]

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:59.232649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:59.232649Z digest=sha256:1e424fda5ba66f5b6d49d2114b012e9948dcebb2b6fdbf34e56489a3fe424ecb

Observation 86d8b7aa-686f-4f73-8bfe-5733b9e62f8e · outbound

This paper cites Kissane, C., Krzyzanowski, R., Conmy, A., and Nanda, N.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Kissane, C., Krzyzanowski, R., Conmy, A., and Nanda, N

Reference 4

Resolution
malformed identifier
no resolver link, observed 2026-08-03T01:14:58.155823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:58.155823Z digest=sha256:1093087877e5560d59c0848edc2f06e099df66e51ddc65165b43837f43054fbe

Observation 4c040916-a8f8-4b11-8907-acb2332be587 · outbound

This paper cites URL https://arxiv.org/abs/2511.08579.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs URL https://arxiv.org/abs/2511.08579

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:58.193388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:58.193388Z digest=sha256:a9a535854605074ae038528dda03df01bc30c9f6eea948fed1f4ce345a44c0ba

Observation 1d899097-a7a3-44c9-8a4f-950fb72b03c9 · outbound

This paper cites Do Large Language Models Latently Perform Multi-Hop Reasoning?.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Do Large Language Models Latently Perform Multi-Hop Reasoning?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:58.578330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:58.578330Z digest=sha256:8e177e0e1208b3f49b91874ea5c28ee409dbee01ecf1e8f5972faabaad6ea52c

Observation 8939c515-2167-4bab-a24c-6d8969e8d3ea · outbound

This paper cites an unresolved cited work.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:58.986640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:58.986640Z digest=sha256:e19d2c5bf9e49cf1273a4e8188aa054ec1b2330d3ba0c9b8fefd17e9109a12e6

Observation 31dd630c-5491-4520-be02-7a8d9d450ca9 · outbound

This paper cites an unresolved cited work.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:59.086880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:59.086880Z digest=sha256:944b363508c568de3ff1ba7c7716b9e3ae1fdeac26532b3029ec9c16ce7d25c0

Observation d03c27a0-fd53-43c6-a345-2f194bd66d8b · outbound

This paper cites an unresolved cited work.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:59.402782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:59.402782Z digest=sha256:ee100e5218ded24fd8a2e8dc2b9b99739b9b30e09765f13b50ed0cee01dbc063

Observation f42e4a61-4db6-42d2-ba5e-24fb6d7322c6 · outbound

This paper cites Left Ginza.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Left Ginza

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:59.524167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:59.524167Z digest=sha256:3c66521f55cace5b325436373340c3d700c127439cb20e0da6b4ef94e79e1d83

Observation 0a16c97a-6e99-4de5-ad89-23360f0470e7 · outbound

This paper cites reasoning.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs reasoning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:59.554865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:59.554865Z digest=sha256:6a632afdc9b3263d3afd770ff361c132a932322d53cb07ad6e0add08776c0714

Observation 2b88c734-a6f3-4eb5-b072-919c81abb261 · outbound

This paper cites Tell me about [topic].

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Tell me about [topic]

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:59.572874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:59.572874Z digest=sha256:5a3a8a510847b3a3989e99d8580b5be3e928c64b150f63ee9956dc7bc2e90d32

Observation 692a6ed7-739b-47f8-928b-f18a41d9c38d · outbound

This paper cites Lindsey, J.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Lindsey, J

Reference 147

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:58.385110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:58.385110Z digest=sha256:429cab0812846fdbb657b9d5d7e68c81d068bf0d8c561d18709c3b61620bccab

Observation 39a3cf3b-d41d-4ac6-aeed-d48ea1255d0c · outbound

This paper cites Towards General Text Embeddings with Multi-stage Contrastive Learning.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Towards General Text Embeddings with Multi-stage Contrastive Learning

Reference 353

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:58.256692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:58.256692Z digest=sha256:2b7aab555cda0774e4da06892b02b1b6e48e4c7c82537be441373383bcf850ac

Observation feddd9b1-463e-42c8-acff-10c4c15954d3 · outbound

This paper cites Llama Scope: Extracting Millions of Features from Llama-3.1-8B with Sparse Autoencoders.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Llama Scope: Extracting Millions of Features from Llama-3.1-8B with Sparse Autoencoders

Reference 600

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:58.085415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:58.085415Z digest=sha256:52b5988b636e74cc8153c73462a9d5260bdf7ee2c4d23a95e824fbd18ce31be2

Observation 2ce183c0-aa24-4abe-93b7-f53a37c0f4f2 · outbound

This paper cites Taboo” baseline uses the following prompt: Describe {topic_phrase} without using the word.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Taboo” baseline uses the following prompt: Describe {topic_phrase} without using the word

Reference 2019

Resolution
malformed identifier
no resolver link, observed 2026-08-03T01:14:59.660318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:59.660318Z digest=sha256:64f19c9bf26845fb5a7f7add360718a1b09b29f12ebf4aa0f54e7a773afc7887

Observation 10837290-1fcd-4e71-b248-61386805d3a6 · outbound

This paper cites cc/paper_files/paper/2022/hash/6f1d4 3d5a82a37e89b0665b33bf3a182-Abstrac t-Conference.html.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs cc/paper_files/paper/2022/hash/6f1d4 3d5a82a37e89b0665b33bf3a182-Abstrac t-Conference.html

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:58.480736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:58.480736Z digest=sha256:dcbf5886fc8b4acd523bca70b4549a6332e1ce7d47c23635774d54005c75ccee

Observation c7cd32bf-f96f-455a-bcdd-f2712dcdd238 · outbound

This paper cites Towards General Text Embeddings with Multi-stage Contrastive Learning.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Towards General Text Embeddings with Multi-stage Contrastive Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:58.325590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:58.325590Z digest=sha256:e04b6d1ceb4c4f448d57827261009a2b9ec9ffa6a696fc5ab1f11fef2ba2fd5a

Observation c574c26f-2878-438c-90e7-dddf62e1d865 · outbound

This paper cites Sparse Autoencoders Find Highly Interpretable Features in Language Models.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Sparse Autoencoders Find Highly Interpretable Features in Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:58.032031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:58.032031Z digest=sha256:95863c64c60afca787075b5e94e99dbb172ab349defb78906121d9b16b519b64

Observation 0fc7b930-0c60-4b34-b99b-1ab1adf79c9f · outbound

This paper cites Eliciting Latent Predictions from Transformers with the Tuned Lens.

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs Eliciting Latent Predictions from Transformers with the Tuned Lens

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T01:14:57.994458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:14:57.994458Z digest=sha256:00b4d6fce52c1d1a54cbcb2a0332dfb227865fd42b9d884c7453db98bade5e42

Pith citing papers

Observation e6bfdad0-2cc0-4030-804c-e618b691e5a1 · inbound

The Grounding Gap: How LLMs Anchor the Meaning of Abstract Concepts Differently from Humans cites this paper.

The Grounding Gap: How LLMs Anchor the Meaning of Abstract Concepts Differently from Humans Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-06-03T02:05:13.472474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-12T03:21:30.932283Z digest=sha256:3ea592166ac6cf9600057a476f0ea8f5d69db810bb38f5f8d933415adb0395e9