Pith. sign in

Paper Citation Record · LEDGER

Towards Unifying Interpretability and Control: Evaluation via Intervention

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2411.04430.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.04430 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T10:46:52.066239Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T22:57:25.994821Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 54b87572-b540-4348-933c-0fa427ff20d5 · inbound

From Attribution to Action: A Human-Centered Application of Activation Steering cites this paper.

From Attribution to Action: A Human-Centered Application of Activation Steering Towards Unifying Interpretability and Control: Evaluation via Intervention

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:21:04.792680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T14:59:12.966608Z digest=sha256:ad3ec75f194574a4aac016df44e047d2dc1ca730f1ebe65fc79bcd872065c509

Observation 8e4084d3-07f5-44a2-8a0a-184cda8de88a · inbound

From Attribution to Action: A Human-Centered Application of Activation Steering cites this paper.

From Attribution to Action: A Human-Centered Application of Activation Steering Towards Unifying Interpretability and Control: Evaluation via Intervention

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T21:57:28.977088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T21:57:28.977088Z digest=sha256:458aa1568a1fa6e9e946b7c7aaf88f894a6d960593278e6a71bd66174b55d41e

Observation b740cc56-a337-49e2-8be4-3055cccbb90b · inbound

SwordBench: Evaluating Orthogonality of Steering Image Representations cites this paper.

SwordBench: Evaluating Orthogonality of Steering Image Representations Towards Unifying Interpretability and Control: Evaluation via Intervention

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:39:09.776150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-20T22:38:24.783740Z digest=sha256:677fee27de25c16345815dbee462898f2fc13b75f43a22f5b0eb01fd5f2b19b9

Observation 56b04572-a408-40f4-8bc9-1d5f45936af1 · inbound

The Shape of Addition: Geometric Structures of Arithmetic in Large Language Models cites this paper.

The Shape of Addition: Geometric Structures of Arithmetic in Large Language Models Towards Unifying Interpretability and Control: Evaluation via Intervention

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:52:49.741820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-28T23:45:16.789926Z digest=sha256:9670470550c5bbee3c5fb0e6683011ec05a1203ce9341acda64a8059138e899a

Observation 2818898b-3b77-4064-9fbe-7524327f9630 · inbound

SAEExplainer: Interpreting SAE Features with Activation-Guided Preference Optimization cites this paper.

SAEExplainer: Interpreting SAE Features with Activation-Guided Preference Optimization Towards Unifying Interpretability and Control: Evaluation via Intervention

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:57:25.996505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-27T18:35:47.513717Z digest=sha256:46c726d293580f41ec11ab1e5c05755212ce49a475985b38c86775d497458655

Observation 548bb33b-35df-4e5c-9e69-c28e8def5b8e · inbound

Position: Explainability Research Must Prioritize Foundations over Ad-hoc Methods cites this paper.

Position: Explainability Research Must Prioritize Foundations over Ad-hoc Methods Towards Unifying Interpretability and Control: Evaluation via Intervention

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T10:46:52.066239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T10:46:52.066239Z digest=sha256:d51080c6fd6be34921e9ef4b20e2514d4f18fec7737fb5d9f048560599ecd91d