Pith. sign in

Paper Citation Record · LEDGER

Improving Activation Steering in Language Models with Mean-Centring

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2312.03813.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.03813 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:49:03.196709Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:27:56.553010Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 64ab4540-b665-49d7-9c5a-dbf3fa146168 · inbound

Refusal in Language Models Is Mediated by a Single Direction cites this paper.

Refusal in Language Models Is Mediated by a Single Direction Improving Activation Steering in Language Models with Mean-Centring

Reference 142

Resolution
verified exact
arxiv_id, observed 2026-05-13T10:47:56.002005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-13T10:47:55.934081Z digest=sha256:d63d30e0b7f899b0d654338328fa00512604616ed5f8e93c1fc363c1869123ce

Observation 2936ed99-97c1-48b5-8fbe-fa8e416647f2 · inbound

Enhancing Semantic Consistency of Large Language Models through Model Editing: An Interpretability-Oriented Approach cites this paper.

Enhancing Semantic Consistency of Large Language Models through Model Editing: An Interpretability-Oriented Approach Improving Activation Steering in Language Models with Mean-Centring

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T18:47:26.159463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T18:47:26.159463Z digest=sha256:c0491240fc5a8fe78805b7dc286c9f4382f80f75876dde5a07409687be12e131

Observation a0e31b78-900a-44e6-983d-9f897c9fa452 · inbound

Risk Assessment Framework for Code LLMs via Leveraging Internal States cites this paper.

Risk Assessment Framework for Code LLMs via Leveraging Internal States Improving Activation Steering in Language Models with Mean-Centring

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T11:49:03.196709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:49:03.196709Z digest=sha256:18f9eaaa4b9d75ec012a11d4900eec0444f12345572c9d32f39f94154fb58d35

Observation 2e96da92-1044-4c69-b58c-6500163e9fcd · inbound

Denoising Concept Vectors with Sparse Autoencoders for Improved Language Model Steering cites this paper.

Denoising Concept Vectors with Sparse Autoencoders for Improved Language Model Steering Improving Activation Steering in Language Models with Mean-Centring

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:30:21.140538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:30:21.140538Z digest=sha256:09cbcc706dc0654214eadf81d4237fc45b20068e3efd80a040ce36bdde9e6200

Observation a89742db-a8b8-4ee6-b2c5-cba54d8426c5 · inbound

Linear Spatial World Models Emerge in Large Language Models cites this paper.

Linear Spatial World Models Emerge in Large Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:27.899636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:27.899636Z digest=sha256:b97577528a97eac8fe033cc36fcf74edd681ae86bd6da1bf16646f77b187f59a

Observation ccda8ba5-f1c8-4351-b338-06c32a8b5839 · inbound

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models cites this paper.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:42.231617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:42.231617Z digest=sha256:35bd52f3d0d695b41bad34d11e8b039678a6c548568874c402420db12b014a5f

Observation cf554321-a7f5-4f89-8968-71a2fe8805f5 · inbound

Probing the Robustness of Large Language Models Safety to Latent Perturbations cites this paper.

Probing the Robustness of Large Language Models Safety to Latent Perturbations Improving Activation Steering in Language Models with Mean-Centring

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:14.924023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:14.924023Z digest=sha256:cb92fda4ae4c13d2e9a0504a4bca05c1195ce8e117927a6593994e5fde527498

Observation b2124474-298c-46db-94ea-6adb117f5cf8 · inbound

From Concepts to Components: Concept-Agnostic Attention Module Discovery in Transformers cites this paper.

From Concepts to Components: Concept-Agnostic Attention Module Discovery in Transformers Improving Activation Steering in Language Models with Mean-Centring

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T19:17:52.270009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:17:52.270009Z digest=sha256:b431ee1c0c425fb8bf9873ff839ed11bd1c08d16a6b4a1075b600a5484a58505

Observation 8258fe95-752b-4992-b837-080c7a43eda1 · inbound

Balancing Stylization and Truth via Disentangled Representation Steering cites this paper.

Balancing Stylization and Truth via Disentangled Representation Steering Improving Activation Steering in Language Models with Mean-Centring

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T23:59:14.682573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:59:14.682573Z digest=sha256:ec06d8ca49671eb1b74753177829b5abf2fb916dc0e0142fd664b15492e00133

Observation eadd3a0b-2ed8-4f62-9c0e-98dafa0b00e2 · inbound

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM cites this paper.

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM Improving Activation Steering in Language Models with Mean-Centring

Reference 162

Resolution
unresolved
no resolver link, observed 2026-08-05T23:13:05.005161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:13:05.005161Z digest=sha256:eaa51e87ac44fa67fa71f6d9089d0eaf013221555eebff9862364418e69f84d4

Observation 1564b451-a3a0-4ac9-80cf-75dad5e35538 · inbound

Multimodal Function Vectors for Visual Relations cites this paper.

Multimodal Function Vectors for Visual Relations Improving Activation Steering in Language Models with Mean-Centring

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T12:44:41.880589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:44:41.880589Z digest=sha256:e89b007717870082f9c57c822e4bf8bb28b355639239a1e9781c7e82ca0109ee

Observation 3eda15f8-7f19-4ffc-a0d3-6c46e23d0ab8 · inbound

Differential syntactic and semantic encoding in LLMs cites this paper.

Differential syntactic and semantic encoding in LLMs Improving Activation Steering in Language Models with Mean-Centring

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T12:03:27.087762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:03:27.087762Z digest=sha256:69e6c8e0783909ee3cd8dbbbd52c58e3ae553d5447e775ab6a113666c2478d79

Observation e956099e-cb86-43b2-aa10-e979664e96df · inbound

The Cylindrical Representation Hypothesis for Language Model Steering cites this paper.

The Cylindrical Representation Hypothesis for Language Model Steering Improving Activation Steering in Language Models with Mean-Centring

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-01T00:15:09.256284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-01T00:10:29.122196Z digest=sha256:9d514b95ad8090126f484820af996a7e80a6aef92e364bb6f8494398b97ac6c2

Observation b5fa2390-e3a1-4908-8ad6-a6fb1ddaa531 · inbound

Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes cites this paper.

Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes Improving Activation Steering in Language Models with Mean-Centring

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:31:10.473529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-08T11:42:16.090162Z digest=sha256:7f5487816777c80a1ae19fa71a860d40921d05b27e31f3245514d2ef0877601f

Observation 59167a3b-6b18-4eae-89d0-5965bd3de790 · inbound

The Granularity Axis: A Micro-to-Macro Latent Direction for Social Roles in Language Models cites this paper.

The Granularity Axis: A Micro-to-Macro Latent Direction for Social Roles in Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:11:09.048343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T10:11:08.494758Z digest=sha256:066f5261a6242bdddc735bfd29bb745d285b9a44dca65657c29f496d89c3adff

Observation 0592810a-e529-472f-aba0-f201b27989b0 · inbound

When is Your LLM Steerable? cites this paper.

When is Your LLM Steerable? Improving Activation Steering in Language Models with Mean-Centring

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:27:56.554416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T10:01:21.682285Z digest=sha256:ca0a89d0f22eaf878c75f5833be4f64f85cc050d5b9bb8bf5cb3e295fc471020

Observation fdca373a-c120-4616-91e0-89efb0207b7a · inbound

Context Is King: How In-Context Specification Shapes the Geometry of Concepts cites this paper.

Context Is King: How In-Context Specification Shapes the Geometry of Concepts Improving Activation Steering in Language Models with Mean-Centring

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T15:17:59.762871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T15:17:59.762871Z digest=sha256:a90d7dbad045b7bf417ca23d2553fd2dd5e5abb3658ae3a2b9dcf9a23bf48b38

Observation 13645654-ef1f-4b39-9d32-d190f42116db · inbound

Where Steering Signals Come From: Activation Source Selection in Activation Steering cites this paper.

Where Steering Signals Come From: Activation Source Selection in Activation Steering Improving Activation Steering in Language Models with Mean-Centring

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T03:00:44.885003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:00:44.885003Z digest=sha256:d51b4e4e9b827d0ab080a0cd10396d223837eb33bbf1879dc479666d60e37921

Observation bb3bfebc-5642-4281-a2a0-d901d3320dfd · inbound

Toward Fine-Grained Forgetting:Attribute Unlearning for Multimodal Large Language Models cites this paper.

Toward Fine-Grained Forgetting:Attribute Unlearning for Multimodal Large Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T00:39:50.627259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:39:50.627259Z digest=sha256:317b53f680fbd3c52daad9fdb6889b8e37fe4a821412732bfebcdf18cd0f3475

Observation 3fddc452-18b3-4819-9178-43a5dec5c4ab · inbound

Subliminal Learning is Non-Semantic Distillation cites this paper.

Subliminal Learning is Non-Semantic Distillation Improving Activation Steering in Language Models with Mean-Centring

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T00:24:40.169417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:24:40.169417Z digest=sha256:0c793b45baa07a67aeb117c6b5eb22566a9d2255905e74be54eeafd0541de4d8