Pith. sign in

Paper Citation Record · LEDGER

Extracting Latent Steering Vectors from Pretrained Language Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2205.05124.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2205.05124 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:05:17.998310Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:27:56.567705Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c2aa1ea0-706e-4b52-a069-b786cb7efcf4 · inbound

Steering Llama 2 via Contrastive Activation Addition cites this paper.

Steering Llama 2 via Contrastive Activation Addition Extracting Latent Steering Vectors from Pretrained Language Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:37:21.234923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T20:37:20.408376Z digest=sha256:7cb288c5a3d8b5b49dd30e624fa795ff0efdd48a4c304576713c0505f6ef17c3

Observation 91217912-0b99-41be-a8ce-fef01a32ef9d · inbound

Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN cites this paper.

Amplify Adjacent Token Differences: Enhancing Long Chain-of-Thought Reasoning with Shift-FFN Extracting Latent Steering Vectors from Pretrained Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:17.998310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:17.998310Z digest=sha256:be2e30c2efc45026be7114cd4317ab36ae1c21dd11f66bc44ae270f5a48b54c5

Observation 774b271f-373e-430d-aa04-f20f3caf04e2 · inbound

Paths Not Taken: Understanding and Mending the Multilingual Factual Recall Pipeline cites this paper.

Paths Not Taken: Understanding and Mending the Multilingual Factual Recall Pipeline Extracting Latent Steering Vectors from Pretrained Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:58:19.396698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:58:19.396698Z digest=sha256:d1430f2c99fb5a14d0c18baebc61a88ff77fa5db098272d6256c8055ab6cea7b

Observation 1f35f1e4-2447-4a26-bde2-41621eebe10a · inbound

Improved Representation Steering for Language Models cites this paper.

Improved Representation Steering for Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:51:31.709484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:51:31.709484Z digest=sha256:da667590c213a9dc8e5679c438c6cbfa53c5ff03b7736b03fdf0c9d42a4f3821

Observation 82e4239a-a60c-4467-ab08-512ba4871739 · inbound

Mind the Quote: Enabling Quotation-Aware Dialogue in LLMs via Plug-and-Play Modules cites this paper.

Mind the Quote: Enabling Quotation-Aware Dialogue in LLMs via Plug-and-Play Modules Extracting Latent Steering Vectors from Pretrained Language Models

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.708567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:35:30.708567Z digest=sha256:ced8d72d702aa835956ff6b43c287d75d17a81ddfac63350826a5180aee6a2a0

Observation 7696cd5b-55a8-4088-b0f2-1a8750d958c1 · inbound

Fine-Grained Interpretation of Political Opinions in Large Language Models cites this paper.

Fine-Grained Interpretation of Political Opinions in Large Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:40:15.757086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:40:15.757086Z digest=sha256:44ddfe26e96e1ccf227dc162e281c1accf6333592384570aa1320215edf9824b

Observation a6d630a0-3cb9-49e3-8836-bfb9274a7ece · inbound

Can Interpretation Predict Behavior on Unseen Data? cites this paper.

Can Interpretation Predict Behavior on Unseen Data? Extracting Latent Steering Vectors from Pretrained Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T19:10:30.409274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:10:30.409274Z digest=sha256:7f71b36032edc5abacd84a9cd4404548d25a131187ad736fe0918d243df9d17d

Observation 724449c9-6fbd-4896-9ea8-2059a456fd76 · inbound

Simple Mechanistic Explanations for Out-Of-Context Reasoning cites this paper.

Simple Mechanistic Explanations for Out-Of-Context Reasoning Extracting Latent Steering Vectors from Pretrained Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:28:44.194532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:28:44.194532Z digest=sha256:37c2f8ba9550a11c5ca9122ddab67f3dad7c5070f7030d0c4881fc999a2ca5f5

Observation 0722ef15-5b41-414c-8a23-32de27c3fb68 · inbound

Quantifying Conversation Drift in MCP via Latent Polytope cites this paper.

Quantifying Conversation Drift in MCP via Latent Polytope Extracting Latent Steering Vectors from Pretrained Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T22:51:28.153771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:51:28.153771Z digest=sha256:1f6b2941d49c3650b4f682d79b91a8e519d40074bbacb4ec822d44ba0127049e

Observation 09fc1d60-2f5c-4ee1-a372-eaa35d99fcd3 · inbound

Steering Protein Language Models cites this paper.

Steering Protein Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:10:12.109738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:10:12.109738Z digest=sha256:4199d0e7b4475c34c7bb1beab4a2ee5d4f24763bbfcf7e889cca89e308369836

Observation 423175fe-4354-4288-a573-ac598925db28 · inbound

Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation cites this paper.

Exploring Language-Agnosticity in Function Vectors: A Case Study in Machine Translation Extracting Latent Steering Vectors from Pretrained Language Models

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T02:53:29.755378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T02:49:55.908142Z digest=sha256:a5126b4667545897a01d4f27a4998fac41592570e1c98d6965d7a71121d571bd

Observation bc2ee5f1-a7f7-445b-b6c2-465bcad80d90 · inbound

Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts cites this paper.

Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts Extracting Latent Steering Vectors from Pretrained Language Models

Reference 127

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:06:07.199909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-09T18:48:04.046370Z digest=sha256:092d3014702948ade88f9a493db7fdf5f7e45ab873885a01bae4fd7ec3bc6a1d

Observation 2bad3404-af5d-479e-bb04-d00a7430df55 · inbound

Steering grids for sparse-autoencoder features: when a top-context label names an activation regime rather than a causal axis cites this paper.

Steering grids for sparse-autoencoder features: when a top-context label names an activation regime rather than a causal axis Extracting Latent Steering Vectors from Pretrained Language Models

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:10:37.052258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T18:52:01.624279Z digest=sha256:ab286aeab943bb9ffee3a2ad3c590a4650e6895300307e253dd285d25f2afc31

Observation b9bb33ba-3b15-4641-8f08-991ce5cdfd33 · inbound

Steering grids for sparse-autoencoder features: when a top-context label names an activation regime rather than a causal axis cites this paper.

Steering grids for sparse-autoencoder features: when a top-context label names an activation regime rather than a causal axis Extracting Latent Steering Vectors from Pretrained Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T14:57:40.631449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:57:40.631449Z digest=sha256:1f4f35aca47a5bbc7c2bc3649d5f6b8a142d3acb5ccc00ee73136628f5c2b4d9

Observation eb27ef93-dcfe-4a19-8f89-71fa180e83f2 · inbound

Manifold Steering Reveals the Shared Geometry of Neural Network Representation and Behavior cites this paper.

Manifold Steering Reveals the Shared Geometry of Neural Network Representation and Behavior Extracting Latent Steering Vectors from Pretrained Language Models

Reference 122

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:16:06.987805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T17:47:09.591001Z digest=sha256:fd4083fac8e969ed20def7fba56a8a1671bc4d5eb44386ad90901ba172cd00c4

Observation bb6281e2-a6f6-44e3-8656-bcda71a4e1e6 · inbound

DataDignity: Training Data Attribution for Large Language Models cites this paper.

DataDignity: Training Data Attribution for Large Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:26:10.297548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-08T11:53:19.594779Z digest=sha256:e3b264e2f39275a271c76806e7a64ea146afc9e02a9c15cd72a732df7cbec144

Observation e5110b5a-2109-43c3-80dd-5407be0851d7 · inbound

Exploitation Without Deception: Dark Triad Feature Steering Reveals Separable Antisocial Circuits in Language Models cites this paper.

Exploitation Without Deception: Dark Triad Feature Steering Reveals Separable Antisocial Circuits in Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:56:19.030279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T02:54:20.313718Z digest=sha256:7aa080b55d011b1bcce80c48da1cf845a9ffe2923a48fd092ed65932b2e14de8

Observation efce1c5c-78c1-400b-87d7-3b3849fa3471 · inbound

Steered Generation via Gradient-Based Optimization on Sparse Query Features cites this paper.

Steered Generation via Gradient-Based Optimization on Sparse Query Features Extracting Latent Steering Vectors from Pretrained Language Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:36:40.329750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T05:31:29.510639Z digest=sha256:e5a9f781e9f8fe3c76c0001f1461e12b1504e5a4fa218c14454411b1e4bb150f

Observation a36bf0d3-956e-4a33-a1a5-6209fcb1ad45 · inbound

When is Your LLM Steerable? cites this paper.

When is Your LLM Steerable? Extracting Latent Steering Vectors from Pretrained Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:27:56.569211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T10:01:21.682285Z digest=sha256:c2e34778d44c64db763ed653e0328296157ab39adf87de1113f5ad3f8551b2c3

Observation 90f6a730-1046-4d1f-9ac6-5a9ea98ac2cd · inbound

Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias cites this paper.

Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias Extracting Latent Steering Vectors from Pretrained Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-14T02:33:34.084111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T02:33:34.084111Z digest=sha256:29209b597a73a28df19b58c20764f76edc3250d51294846be80a8f559c230668

Observation 71d3ad5d-1205-48ce-8f08-21638e79ba52 · inbound

Probabilistic Concept-Aware Steering for Trustworthy LLM Inference cites this paper.

Probabilistic Concept-Aware Steering for Trustworthy LLM Inference Extracting Latent Steering Vectors from Pretrained Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T13:56:44.575890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T13:56:44.575890Z digest=sha256:623bc340ee305b717478340554fd1bb511ac52e7e6a6cd37a9df21479bf667a6