Pith. sign in

Paper Citation Record · LEDGER

Improving Activation Steering in Language Models with Mean-Centring

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2312.03813.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.03813 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:30:21.140538Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:27:56.553010Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 64ab4540-b665-49d7-9c5a-dbf3fa146168 · inbound

Refusal in Language Models Is Mediated by a Single Direction cites this paper.

Refusal in Language Models Is Mediated by a Single Direction Improving Activation Steering in Language Models with Mean-Centring

Reference 142

Resolution
verified exact
arxiv_id, observed 2026-05-13T10:47:56.002005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T10:47:55.934081Z digest=sha256:ad1a086284260c7d44de0ab3bced308c9adb22fcb87691f2a7ccf239cab8ee26

Observation 2e96da92-1044-4c69-b58c-6500163e9fcd · inbound

Denoising Concept Vectors with Sparse Autoencoders for Improved Language Model Steering cites this paper.

Denoising Concept Vectors with Sparse Autoencoders for Improved Language Model Steering Improving Activation Steering in Language Models with Mean-Centring

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:30:21.140538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:30:21.140538Z digest=sha256:9143604aa0cc08806b96184dd31bb1ca25f7452c5c77f4b8fe127f0970817143

Observation a89742db-a8b8-4ee6-b2c5-cba54d8426c5 · inbound

Linear Spatial World Models Emerge in Large Language Models cites this paper.

Linear Spatial World Models Emerge in Large Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:17:27.899636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:17:27.899636Z digest=sha256:2838501b6147ef4de30bdacccbff2b8104f9c0028146a6ad0640b00378f41cf5

Observation ccda8ba5-f1c8-4351-b338-06c32a8b5839 · inbound

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models cites this paper.

From Emergence to Control: Probing and Modulating Self-Reflection in Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T01:03:42.231617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:03:42.231617Z digest=sha256:d426b2450b0da29dc3f68447632107c63a460ddfb6623d871bdbd92519d5a32b

Observation cf554321-a7f5-4f89-8968-71a2fe8805f5 · inbound

Probing the Robustness of Large Language Models Safety to Latent Perturbations cites this paper.

Probing the Robustness of Large Language Models Safety to Latent Perturbations Improving Activation Steering in Language Models with Mean-Centring

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:14.924023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:14.924023Z digest=sha256:5728dc9a57ce2931679b8eaf87b84ce9c2c5301edb0fadcd8f17a9217be11fea

Observation 8258fe95-752b-4992-b837-080c7a43eda1 · inbound

Balancing Stylization and Truth via Disentangled Representation Steering cites this paper.

Balancing Stylization and Truth via Disentangled Representation Steering Improving Activation Steering in Language Models with Mean-Centring

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T23:59:14.682573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:59:14.682573Z digest=sha256:bfeedd0409d9e091e217eec33e1c03544afc87deda0207cd6cfde7dd9d6e4e77

Observation eadd3a0b-2ed8-4f62-9c0e-98dafa0b00e2 · inbound

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM cites this paper.

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM Improving Activation Steering in Language Models with Mean-Centring

Reference 162

Resolution
unresolved
no resolver link, observed 2026-08-05T23:13:05.005161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:13:05.005161Z digest=sha256:1af1b9d41af871558bcc54a7c6912bd88b2d68ca428c4a680dd80d3558165738

Observation 1564b451-a3a0-4ac9-80cf-75dad5e35538 · inbound

Multimodal Function Vectors for Visual Relations cites this paper.

Multimodal Function Vectors for Visual Relations Improving Activation Steering in Language Models with Mean-Centring

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T12:44:41.880589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:44:41.880589Z digest=sha256:fcb9712a4a117ae81bf1dbf173e8b5d538525837c57664b31b00b8b238af16fc

Observation 3eda15f8-7f19-4ffc-a0d3-6c46e23d0ab8 · inbound

Differential syntactic and semantic encoding in LLMs cites this paper.

Differential syntactic and semantic encoding in LLMs Improving Activation Steering in Language Models with Mean-Centring

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T12:03:27.087762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:03:27.087762Z digest=sha256:559d7c61303aff72b3c8b60567d62e12e5c0e8ca5562dd5cf143efd5a619281a

Observation e956099e-cb86-43b2-aa10-e979664e96df · inbound

The Cylindrical Representation Hypothesis for Language Model Steering cites this paper.

The Cylindrical Representation Hypothesis for Language Model Steering Improving Activation Steering in Language Models with Mean-Centring

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-01T00:15:09.256284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T00:10:29.122196Z digest=sha256:a034435178be375658d7e365f70c0ab7f9d4e5f70c99941ae2cd6bffbffd0323

Observation b5fa2390-e3a1-4908-8ad6-a6fb1ddaa531 · inbound

Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes cites this paper.

Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes Improving Activation Steering in Language Models with Mean-Centring

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:31:10.473529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T11:42:16.090162Z digest=sha256:22aa14222d439af4b8513b5b13b976e162e593e095565551bb8078be3d9151d2

Observation 59167a3b-6b18-4eae-89d0-5965bd3de790 · inbound

The Granularity Axis: A Micro-to-Macro Latent Direction for Social Roles in Language Models cites this paper.

The Granularity Axis: A Micro-to-Macro Latent Direction for Social Roles in Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:11:09.048343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T10:11:08.494758Z digest=sha256:39b9306174e594b8cc2f4f78059c54996273dfddf80afad9877b5cefe0718180

Observation 0592810a-e529-472f-aba0-f201b27989b0 · inbound

When is Your LLM Steerable? cites this paper.

When is Your LLM Steerable? Improving Activation Steering in Language Models with Mean-Centring

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:27:56.554416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T10:01:21.682285Z digest=sha256:5ac46476618d2f51a56c48622d4a18b4e4abbb9a6714ea7e07ce9f313187290c

Observation fdca373a-c120-4616-91e0-89efb0207b7a · inbound

Context Is King: How In-Context Specification Shapes the Geometry of Concepts cites this paper.

Context Is King: How In-Context Specification Shapes the Geometry of Concepts Improving Activation Steering in Language Models with Mean-Centring

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-31T15:17:59.762871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T15:17:59.762871Z digest=sha256:d8e6ade05d5e480d71819913c4ff41dadc9ab49a35b7071178c41cb021ef7fc1

Observation 13645654-ef1f-4b39-9d32-d190f42116db · inbound

Where Steering Signals Come From: Activation Source Selection in Activation Steering cites this paper.

Where Steering Signals Come From: Activation Source Selection in Activation Steering Improving Activation Steering in Language Models with Mean-Centring

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T03:00:44.885003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:00:44.885003Z digest=sha256:7a02625c6196e62786663389d99d223a6432c27a5ff78f1efc1ea353cd98a10e

Observation bb3bfebc-5642-4281-a2a0-d901d3320dfd · inbound

Toward Fine-Grained Forgetting:Attribute Unlearning for Multimodal Large Language Models cites this paper.

Toward Fine-Grained Forgetting:Attribute Unlearning for Multimodal Large Language Models Improving Activation Steering in Language Models with Mean-Centring

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T00:39:50.627259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:39:50.627259Z digest=sha256:862ecf37fdec3f3611288cb073d4aa5f897958534ab866bc071a566162f6147e