Pith. sign in

Paper Citation Record · LEDGER

Extracting Latent Steering Vectors from Pretrained Language Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2205.05124.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2205.05124 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:05:17.998310Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:27:56.567705Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c2aa1ea0-706e-4b52-a069-b786cb7efcf4 · inbound

Steering Llama 2 via Contrastive Activation Addition cites this paper.

Steering Llama 2 via Contrastive Activation Addition Extracting Latent Steering Vectors from Pretrained Language Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:37:21.234923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-11T20:37:20.408376Z digest=sha256:a4aab116f214cab633dacbccff38267c0db3c7d69d9228cce23c942114b14c9d

Observation 91217912-0b99-41be-a8ce-fef01a32ef9d · inbound

Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN cites this paper.

Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Extracting Latent Steering Vectors from Pretrained Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:17.998310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:17.998310Z digest=sha256:be2e30c2efc45026be7114cd4317ab36ae1c21dd11f66bc44ae270f5a48b54c5

Observation 774b271f-373e-430d-aa04-f20f3caf04e2 · inbound

Paths Not Taken: Understanding and Mending the Multilingual Factual Recall Pipeline cites this paper.

Paths Not Taken: Understanding and Mending the Multilingual Factual Recall Pipeline Extracting Latent Steering Vectors from Pretrained Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:58:19.396698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:58:19.396698Z digest=sha256:d1430f2c99fb5a14d0c18baebc61a88ff77fa5db098272d6256c8055ab6cea7b

Observation 1f35f1e4-2447-4a26-bde2-41621eebe10a · inbound

Improved Representation Steering for Language Models cites this paper.

Improved Representation Steering for Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:51:31.709484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:51:31.709484Z digest=sha256:da667590c213a9dc8e5679c438c6cbfa53c5ff03b7736b03fdf0c9d42a4f3821

Observation 82e4239a-a60c-4467-ab08-512ba4871739 · inbound

Mind the Quote: Enabling Quotation-Aware Dialogue in LLMs via Plug-and-Play Modules cites this paper.

Mind the Quote: Enabling Quotation-Aware Dialogue in LLMs via Plug-and-Play Modules Extracting Latent Steering Vectors from Pretrained Language Models

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.708567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:35:30.708567Z digest=sha256:ced8d72d702aa835956ff6b43c287d75d17a81ddfac63350826a5180aee6a2a0

Observation 7696cd5b-55a8-4088-b0f2-1a8750d958c1 · inbound

Fine-Grained Interpretation of Political Opinions in Large Language Models cites this paper.

Fine-Grained Interpretation of Political Opinions in Large Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:40:15.757086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:40:15.757086Z digest=sha256:44ddfe26e96e1ccf227dc162e281c1accf6333592384570aa1320215edf9824b

Observation a6d630a0-3cb9-49e3-8836-bfb9274a7ece · inbound

Can Interpretation Predict Behavior on Unseen Data? cites this paper.

Can Interpretation Predict Behavior on Unseen Data? Extracting Latent Steering Vectors from Pretrained Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T19:10:30.409274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:10:30.409274Z digest=sha256:7f71b36032edc5abacd84a9cd4404548d25a131187ad736fe0918d243df9d17d

Observation 724449c9-6fbd-4896-9ea8-2059a456fd76 · inbound

Simple Mechanistic Explanations for Out-Of-Context Reasoning cites this paper.

Simple Mechanistic Explanations for Out-Of-Context Reasoning Extracting Latent Steering Vectors from Pretrained Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:28:44.194532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:28:44.194532Z digest=sha256:37c2f8ba9550a11c5ca9122ddab67f3dad7c5070f7030d0c4881fc999a2ca5f5

Observation 0722ef15-5b41-414c-8a23-32de27c3fb68 · inbound

Quantifying Conversation Drift in MCP via Latent Polytope cites this paper.

Quantifying Conversation Drift in MCP via Latent Polytope Extracting Latent Steering Vectors from Pretrained Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T22:51:28.153771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:51:28.153771Z digest=sha256:1f6b2941d49c3650b4f682d79b91a8e519d40074bbacb4ec822d44ba0127049e

Observation 09fc1d60-2f5c-4ee1-a372-eaa35d99fcd3 · inbound

Steering Protein Language Models cites this paper.

Steering Protein Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:12.109738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:12.109738Z digest=sha256:4199d0e7b4475c34c7bb1beab4a2ee5d4f24763bbfcf7e889cca89e308369836

Observation 423175fe-4354-4288-a573-ac598925db28 · inbound

Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation cites this paper.

Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation Extracting Latent Steering Vectors from Pretrained Language Models

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T02:53:29.755378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T02:49:55.908142Z digest=sha256:9fc961e62b7987f8105b80526b5710f8068ce75abbf9dd206610643a0e0f2a0f

Observation bc2ee5f1-a7f7-445b-b6c2-465bcad80d90 · inbound

Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts cites this paper.

Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts Extracting Latent Steering Vectors from Pretrained Language Models

Reference 127

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:06:07.199909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-09T18:48:04.046370Z digest=sha256:b15fd916470b549319121552e6ceeedccdcf4646a99e8c04a5a7ed4ea3525859

Observation 2bad3404-af5d-479e-bb04-d00a7430df55 · inbound

Steering grids for sparse-autoencoder features: when a top-context label names an activation regime rather than a causal axis cites this paper.

Steering grids for sparse-autoencoder features: when a top-context label names an activation regime rather than a causal axis Extracting Latent Steering Vectors from Pretrained Language Models

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:10:37.052258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T18:52:01.624279Z digest=sha256:11c5a7fd5d7c34341451e2b03be12f5f653c811adb72eba48b5aad2c2e137d2f

Observation b9bb33ba-3b15-4641-8f08-991ce5cdfd33 · inbound

Steering grids for sparse-autoencoder features: when a top-context label names an activation regime rather than a causal axis cites this paper.

Steering grids for sparse-autoencoder features: when a top-context label names an activation regime rather than a causal axis Extracting Latent Steering Vectors from Pretrained Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T14:57:40.631449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:57:40.631449Z digest=sha256:1f4f35aca47a5bbc7c2bc3649d5f6b8a142d3acb5ccc00ee73136628f5c2b4d9

Observation eb27ef93-dcfe-4a19-8f89-71fa180e83f2 · inbound

Manifold Steering Reveals the Shared Geometry of Neural Network Representation and Behavior cites this paper.

Manifold Steering Reveals the Shared Geometry of Neural Network Representation and Behavior Extracting Latent Steering Vectors from Pretrained Language Models

Reference 122

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:16:06.987805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T17:47:09.591001Z digest=sha256:5c847c660342d4ca5d1984ceb61629f4aeb828d3f30e9c4e79a087ecf57e8bfc

Observation bb6281e2-a6f6-44e3-8656-bcda71a4e1e6 · inbound

DataDignity: Training Data Attribution for Large Language Models cites this paper.

DataDignity: Training Data Attribution for Large Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:26:10.297548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T11:53:19.594779Z digest=sha256:25190058edf4f434423afa90ad3c3f7f05270ba1b47aae31965ac8b1dee417c5

Observation e5110b5a-2109-43c3-80dd-5407be0851d7 · inbound

Exploitation Without Deception: Dark Triad Feature Steering Reveals Separable Antisocial Circuits in Language Models cites this paper.

Exploitation Without Deception: Dark Triad Feature Steering Reveals Separable Antisocial Circuits in Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:56:19.030279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T02:54:20.313718Z digest=sha256:2744fabcba93e8a5a53952d26edaf2f6e5e7071bad315468cf8a562b4a6670be

Observation efce1c5c-78c1-400b-87d7-3b3849fa3471 · inbound

Steered Generation via Gradient-Based Optimization on Sparse Query Features cites this paper.

Steered Generation via Gradient-Based Optimization on Sparse Query Features Extracting Latent Steering Vectors from Pretrained Language Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:36:40.329750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T05:31:29.510639Z digest=sha256:695e95db46d117ab35e7dd430d00f2fba4610ac3c927480ce0308c230b818a24

Observation a36bf0d3-956e-4a33-a1a5-6209fcb1ad45 · inbound

When is Your LLM Steerable? cites this paper.

When is Your LLM Steerable? Extracting Latent Steering Vectors from Pretrained Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:27:56.569211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T10:01:21.682285Z digest=sha256:19de202730db7379f88351cbf1e13d75505a585f0c6ab42f88c08597b68b2b34

Observation 90f6a730-1046-4d1f-9ac6-5a9ea98ac2cd · inbound

Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias cites this paper.

Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias Extracting Latent Steering Vectors from Pretrained Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-14T02:33:34.084111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T02:33:34.084111Z digest=sha256:29209b597a73a28df19b58c20764f76edc3250d51294846be80a8f559c230668

Observation 71d3ad5d-1205-48ce-8f08-21638e79ba52 · inbound

Probabilistic Concept-Aware Steering for Trustworthy LLM Inference cites this paper.

Probabilistic Concept-Aware Steering for Trustworthy LLM Inference Extracting Latent Steering Vectors from Pretrained Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T13:56:44.575890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T13:56:44.575890Z digest=sha256:623bc340ee305b717478340554fd1bb511ac52e7e6a6cd37a9df21479bf667a6