Pith. sign in

Paper Citation Record · LEDGER

Understanding (Un)Reliability of Steering Vectors in Language Models

As of 7 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 10 inbound Pith citation observations for arXiv:2505.22637.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22637 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:10:08.895546Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T11:08:17.571722Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T15:49:57.002506Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact2
  • verified fuzzy6
  • unresolved28
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bd450999-42e0-4f54-877d-862e670f8720 · outbound

This paper cites write newline.

Understanding (Un)Reliability of Steering Vectors in Language Models write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:03.289827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:03.289827Z digest=sha256:cade30b52d47d4a23dd69ec2ddb029acf15df77719d28089c2b2bbfc080a91ac

Observation b12abee5-3591-465f-a777-29e583a26be2 · outbound

This paper cites Refusal in Language Models Is Mediated by a Single Direction.

Understanding (Un)Reliability of Steering Vectors in Language Models Refusal in Language Models Is Mediated by a Single Direction

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:03.419878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:03.419878Z digest=sha256:5f43c4e92922331e9ecb000a6c5b6f5a883742f6e954e52b3d675250c1e504ee

Observation 4955598a-7f3d-400c-ae32-003701625648 · outbound

This paper cites Cats: Customizable abstractive topic-based summarization.

Understanding (Un)Reliability of Steering Vectors in Language Models Cats: Customizable abstractive topic-based summarization

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:03.548101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:03.548101Z digest=sha256:d5415ac21e6f9c6c12ad4fd580ce0435a38f2dd511a018e0bfbcdee173e22584

Observation 6d857f70-6587-4148-a88f-847eab514349 · outbound

This paper cites NEWTS : A corpus for news topic-focused summarization.

Understanding (Un)Reliability of Steering Vectors in Language Models NEWTS : A corpus for news topic-focused summarization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:03.715918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:03.715918Z digest=sha256:c094566dbe26a2699f87d8038a06b6f710e0131df7d9552d4d4b1343a1a80e76

Observation 8ae392d8-655e-4c49-a154-3e3524d501bf · outbound

This paper cites Controllable Topic-Focused Abstractive Summarization.

Understanding (Un)Reliability of Steering Vectors in Language Models Controllable Topic-Focused Abstractive Summarization

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:03.927972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:03.927972Z digest=sha256:59691e456a73a70b1bc2d55bb5e0c09838bb009a0f10ec960c950b269b653b82

Observation 7062c719-3fd9-49f0-ac2d-cf6212c10207 · outbound

This paper cites Text simplification via adaptive teaching.

Understanding (Un)Reliability of Steering Vectors in Language Models Text simplification via adaptive teaching

Reference 6

Resolution
verified exact
doi, observed 2026-08-07T13:10:09.638480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:10:04.074437Z digest=sha256:aab8bff8edf055ce37e5b60cf5ee7197f7f111d9dafe48170cd153b1601304f4

Observation e5da340c-198b-477c-b853-a1b095413748 · outbound

This paper cites SIMSUM : Document-level text simplification via simultaneous summarization.

Understanding (Un)Reliability of Steering Vectors in Language Models SIMSUM : Document-level text simplification via simultaneous summarization

Reference 7

Resolution
verified exact
doi, observed 2026-08-07T13:10:09.260944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:10:04.258231Z digest=sha256:6aa95ff5232bc6755310377fc8272d9a6d723786fcb0a2a83bcf71d14555a32a

Observation 2a3eaa2a-7361-4b48-a088-8e8edad3ab5a · outbound

This paper cites A sober look at steering vectors for llms.

Understanding (Un)Reliability of Steering Vectors in Language Models A sober look at steering vectors for llms

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:12.329122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:10:04.453529Z digest=sha256:35d0eb4d5608232d724ae59cba03d0a4c8d2025309d28ec9af9961158170764f

Observation 7110b613-adc6-48fa-a5cd-ec86b0bfb4b0 · outbound

This paper cites Comparing Bottom-Up and Top-Down Steering Approaches on In-Context Learning Tasks.

Understanding (Un)Reliability of Steering Vectors in Language Models Comparing Bottom-Up and Top-Down Steering Approaches on In-Context Learning Tasks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.575056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:04.575056Z digest=sha256:4b9feb3dbd54c1a6d5873458127a1790470bf6e31b7d9ba90df3351f2da67dde

Observation 30fa9f24-32c4-46c4-b4aa-88e21828886b · outbound

This paper cites Personalized Steering of Large Language Models: Versatile Steering Vectors Through Bi-directional Preference Optimization.

Understanding (Un)Reliability of Steering Vectors in Language Models Personalized Steering of Large Language Models: Versatile Steering Vectors Through Bi-directional Preference Optimization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.717862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:04.717862Z digest=sha256:62cf43eaafbae5670d19f7b0ce54370d5adb456fae3d4c64f12b4be7bde607ff

Observation 3701d840-1f93-448f-a62d-ef5ace6f7da4 · outbound

This paper cites In-context learning creates task vectors.

Understanding (Un)Reliability of Steering Vectors in Language Models In-context learning creates task vectors

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.898405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:04.898405Z digest=sha256:b6d887a2f442bd865f7dc4b1d78112a065e367fce6d1f8f686b56b2c335b15fe

Observation 11affa31-d0f5-42b0-847d-02f596676574 · outbound

This paper cites Measuring massive multitask language understanding.

Understanding (Un)Reliability of Steering Vectors in Language Models Measuring massive multitask language understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:04.974038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:04.974038Z digest=sha256:7c2c2d736fe0278a353a95f2df4e4341a0cdc12c08262a7455c9ea43caa4aee1

Observation 87fc6599-bc67-431b-ba03-3660781afacb · outbound

This paper cites Style Vectors for Steering Generative Large Language Models.

Understanding (Un)Reliability of Steering Vectors in Language Models Style Vectors for Steering Generative Large Language Models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:11.949019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:10:05.093975Z digest=sha256:9c65a5671effffb7863d48471a438e15a9804aa6a2f17bfd760994930cf7ff90

Observation 514b7c2d-d754-4bc9-8e5e-7bd005f009f9 · outbound

This paper cites Steering clear: A systematic study of activation steering in a toy setup.

Understanding (Un)Reliability of Steering Vectors in Language Models Steering clear: A systematic study of activation steering in a toy setup

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:11.664519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:10:05.218312Z digest=sha256:dae28f79f3318bab41ee6f0b9e5785dda4c3a9d0df45f50b6c4217b6b8ad6995

Observation 7da431f6-36ab-49cf-a44e-f9ce9e74cf2a · outbound

This paper cites Inference-time intervention: Eliciting truthful answers from a language model.

Understanding (Un)Reliability of Steering Vectors in Language Models Inference-time intervention: Eliciting truthful answers from a language model

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:05.387255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:05.387255Z digest=sha256:673f790bff7de76aa0c78b6f844899c85837e081d8aed700c666689eeaf25466

Observation 49a3bbf8-7c4b-4cd6-bbe5-b215a7fb408f · outbound

This paper cites The Geometry of Truth: Emergent Linear Structure in Large Language Model Representations of True/False Datasets.

Understanding (Un)Reliability of Steering Vectors in Language Models The Geometry of Truth: Emergent Linear Structure in Large Language Model Representations of True/False Datasets

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:05.520729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:05.520729Z digest=sha256:2f8dfe706e4464df7e3fe919d92731d4746fce996770bee3f05747113c0c56c4

Observation 58eb0b15-9f5f-4252-a074-fa17f2772e13 · outbound

This paper cites The geometry of truth: Emergent linear structure in large language model representations of true/false datasets.

Understanding (Un)Reliability of Steering Vectors in Language Models The geometry of truth: Emergent linear structure in large language model representations of true/false datasets

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:11.347820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:10:05.648088Z digest=sha256:8c96967ea4923aa48f6fc1debc354dea4d9f394e7ee47b6fe5c820715ef64c5e

Observation 0bc45b64-edcd-4965-a435-cc549a2067eb · outbound

This paper cites Refusal in LLMs is an Affine Function.

Understanding (Un)Reliability of Steering Vectors in Language Models Refusal in LLMs is an Affine Function

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:05.846705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:05.846705Z digest=sha256:b3a9c7eca7a83beb3d49dc49f8fab76502d5ac272870c29af15ca1f056b24a05

Observation 89d8d63f-d64f-4a5e-93a8-1b2ced9f909a · outbound

This paper cites Towards Reliable Evaluation of Behavior Steering Interventions in LLMs.

Understanding (Un)Reliability of Steering Vectors in Language Models Towards Reliable Evaluation of Behavior Steering Interventions in LLMs

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:06.057729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:06.057729Z digest=sha256:07461d1395e054b53bbdb878059aaeb860a9330a64b4808ede1b553d4a528141

Observation 116acc5e-5900-466e-bca0-86616e6a68de · outbound

This paper cites Steering llama 2 via contrastive activation addition.

Understanding (Un)Reliability of Steering Vectors in Language Models Steering llama 2 via contrastive activation addition

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:06.173597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:06.173597Z digest=sha256:f7169c9ce12e0c07a0820df697eb2b29d34a4839eae97ed01eaaedcf0c5e9f98

Observation dad2db02-e247-492c-b995-e6a20608a38d · outbound

This paper cites Discovering Language Model Behaviors with Model-Written Evaluations, Toronto, Canada, July 2023.

Understanding (Un)Reliability of Steering Vectors in Language Models Discovering Language Model Behaviors with Model-Written Evaluations, Toronto, Canada, July 2023

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:06.293291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:06.293291Z digest=sha256:94f576c6d6fdd14bb5ff81ba7c7928baf8229c206527ec72f1980ecc38f61d6c

Observation 83e09f5f-6664-46c3-bb9c-07f5ade319fd · outbound

This paper cites Representation surgery: Theory and practice of affine steering.

Understanding (Un)Reliability of Steering Vectors in Language Models Representation surgery: Theory and practice of affine steering

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:10.985849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:10:06.461530Z digest=sha256:7abfac3e3d7dbc30ccf60b45c73b46b5db913df99dafc50e4938c334e1172a94

Observation e234c100-7387-4af3-943b-c7b0bbf7c2df · outbound

This paper cites an unresolved cited work.

Understanding (Un)Reliability of Steering Vectors in Language Models Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:10.614343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:10:06.605117Z digest=sha256:c5bb8d30030c2ade6fb5014100654be6d38cf95d4c19e9830a896c9912c9572f

Observation df5ab5eb-3fdb-4ef0-95a3-796919775ffb · outbound

This paper cites Extracting Latent Steering Vectors from Pretrained Language Models.

Understanding (Un)Reliability of Steering Vectors in Language Models Extracting Latent Steering Vectors from Pretrained Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:06.744243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:06.744243Z digest=sha256:5316c7e2492084a4219e8f40a3599957d0fee6c15cc6e8f6ffb00060cd1fd7e9

Observation e27cad75-e919-4e0a-968b-829da0f359a5 · outbound

This paper cites Analysing the generalisation and reliability of steering vectors.

Understanding (Un)Reliability of Steering Vectors in Language Models Analysing the generalisation and reliability of steering vectors

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:06.928187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:06.928187Z digest=sha256:8848d61aa415c348a23becf9df4c3a17d353e1f23b2978cfd29e2c6d949d09d6

Observation 012e9b4e-7a8b-4feb-b47d-7865abd582f6 · outbound

This paper cites Linear Representations of Sentiment in Large Language Models.

Understanding (Un)Reliability of Steering Vectors in Language Models Linear Representations of Sentiment in Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:07.072130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:07.072130Z digest=sha256:ac10e895a62cdc7d7ce57f3b10a3748d1dac54a3183e9903875516a847cb49ae

Observation 0b1763f6-acb5-41dc-871c-d36007e80b69 · outbound

This paper cites Hollinsworth, Atticus Geiger, and Neel Nanda.

Understanding (Un)Reliability of Steering Vectors in Language Models Hollinsworth, Atticus Geiger, and Neel Nanda

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:07.202080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:07.202080Z digest=sha256:70c055ec4875a20eb2e5611e6a8f29ede835e05c5f0adb75f56df43e442f24ec

Observation 96dbf721-bcbc-4538-8d96-a39327a76187 · outbound

This paper cites Function Vectors in Large Language Models.

Understanding (Un)Reliability of Steering Vectors in Language Models Function Vectors in Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:07.335261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:07.335261Z digest=sha256:a8b0bd90643605a33972faac83b62bad89f838d54156e16adb4edf467e54736a

Observation 1608ff49-3205-4878-9d8d-633a7ba8e9d4 · outbound

This paper cites Function vectors in large language models.

Understanding (Un)Reliability of Steering Vectors in Language Models Function vectors in large language models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:07.486495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:07.486495Z digest=sha256:9158e3759531935dddef7c274e2416780e093b39f164cc31dc9f252fa64f818e

Observation 71d1e75e-8ae7-4697-9eaf-bf08c20d67c3 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Understanding (Un)Reliability of Steering Vectors in Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:07.663220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:07.663220Z digest=sha256:c73b40b60f3a27e7920c69f64510f245b395b0d6f299f9b25d5d1a8fa3e29b98

Observation f0ddce0c-bb20-4bec-a1c4-608a2c6799e3 · outbound

This paper cites Activation addition: Steering language models without optimization.

Understanding (Un)Reliability of Steering Vectors in Language Models Activation addition: Steering language models without optimization

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:10.366568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:10:07.830192Z digest=sha256:634ded646446d479fd5bf301b90480e126701f3477947d313a08a9c949deec02

Observation b11d6994-d56f-4cb9-ba92-94582b225ee5 · outbound

This paper cites Controllable text summarization: Unraveling challenges, approaches, and prospects - a survey.

Understanding (Un)Reliability of Steering Vectors in Language Models Controllable text summarization: Unraveling challenges, approaches, and prospects - a survey

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:07.935345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:07.935345Z digest=sha256:f18f3fe0abb46e06657ab2028357227a39c7bad38cd212c822d2896e16704dc4

Observation e52a79ce-8217-4802-8fe8-e87911df09bb · outbound

This paper cites A comprehensive survey on process-oriented automatic text summarization with exploration of llm-based methods, 2025.

Understanding (Un)Reliability of Steering Vectors in Language Models A comprehensive survey on process-oriented automatic text summarization with exploration of llm-based methods, 2025

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:08.060573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:08.060573Z digest=sha256:fd960f1e410206e7eccae02e3f23914d8cea1c9812fbd4e849a5e153279c7d0d

Observation 30e6f97d-3aac-4854-8e10-5df3e4d72504 · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

Understanding (Un)Reliability of Steering Vectors in Language Models Representation Engineering: A Top-Down Approach to AI Transparency

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:08.399853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:08.399853Z digest=sha256:1415f2eb001e6fda707515e0e254d26d95a0afab233ac2bbf1e0fba6ae74f427

Observation 8b8eca33-f2ad-4632-8d6c-55f6e08636e5 · outbound

This paper cites @esa (Ref.

Understanding (Un)Reliability of Steering Vectors in Language Models @esa (Ref

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:08.563640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:08.563640Z digest=sha256:c199ad76d01d653bf9d41c9c405df5a6594e1be4aaa9c890b5eee68be5d4f47c

Observation 400ebaa9-c888-45ae-ae67-5b09f2f630c2 · outbound

This paper cites an unresolved cited work.

Understanding (Un)Reliability of Steering Vectors in Language Models Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:08.737088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:08.737088Z digest=sha256:08f329361b68eed28e27b12afe3e4888b66f0693c697fdbd7f2c98601561b284

Observation d56fbc92-ed8d-4e88-b60b-b678e757acee · outbound

This paper cites 3!( 4˜ "3!( 4˒.

Understanding (Un)Reliability of Steering Vectors in Language Models 3!( 4˜ "3!( 4˒

Reference 38

Resolution
malformed identifier
no resolver link, observed 2026-08-07T13:10:08.895546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:08.895546Z digest=sha256:7764bbbb2b699d205108b16faee3fba9fbd5a1f5cafc20c83f8368e7b80efc36

Pith citing papers

Observation 70274d98-43c8-4562-bdd8-b51bc6ccbdfe · inbound

RouteHijack: Routing-Aware Attack on Mixture-of-Experts LLMs cites this paper.

RouteHijack: Routing-Aware Attack on Mixture-of-Experts LLMs Understanding (Un)Reliability of Steering Vectors in Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:46:13.825157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-09T19:22:00.217729Z digest=sha256:6c6adf8a3b63097ca94b020b80940c02ada8cb0e2b77431b2aa96b5f7f6dc021

Observation 69162c32-1c61-409f-b86e-267231d7d9cd · inbound

Prompt-Activation Duality: Improving Activation Steering via Attention-Level Interventions cites this paper.

Prompt-Activation Duality: Improving Activation Steering via Attention-Level Interventions Understanding (Un)Reliability of Steering Vectors in Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:16:25.939080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T04:27:15.490694Z digest=sha256:ee11c15ee464e250b687ec85c08fa308ff2b76309ae0ff1c43c4b39b3bef9650

Observation 71c24dca-f1f9-44a7-b847-9c1a5afd1c8e · inbound

Prompt-Activation Duality: Improving Activation Steering via Attention-Level Interventions cites this paper.

Prompt-Activation Duality: Improving Activation Steering via Attention-Level Interventions Understanding (Un)Reliability of Steering Vectors in Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:29:47.076007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T05:27:59.011749Z digest=sha256:0fce29a92210b900cc1cacfa36c310e45ca3bbcce9f5af1cfcee30c06351a30a

Observation 1e77b7e1-96eb-4d6f-92e0-237f38dca4b5 · inbound

When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search cites this paper.

When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search Understanding (Un)Reliability of Steering Vectors in Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:39:10.703623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T22:36:12.466550Z digest=sha256:370d301c675e79c4721dd20df429d940c3ba65562f6448e9cfde11d0dfcd3227

Observation b48fd55b-ee3c-4512-a253-0d40a3b82cbb · inbound

When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search cites this paper.

When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search Understanding (Un)Reliability of Steering Vectors in Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:01:23.525559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T09:59:32.201925Z digest=sha256:29b55e5bf25e25b779a0159dc4eb3a1d42f44911115b2b1710a850c01517e67f

Observation 3f421cf8-48ca-41b2-a447-55420c0c24c7 · inbound

Temporal Preference Concepts and their Functions in a Large Language Model cites this paper.

Temporal Preference Concepts and their Functions in a Large Language Model Understanding (Un)Reliability of Steering Vectors in Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:05:47.215257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T22:16:47.743387Z digest=sha256:e78f1fe665b68ce027d6b3d5ebffca9ebaf595c06a1844e384c0b6c1989b82f9

Observation 19796388-a601-4242-a1c8-02fb3a196a44 · inbound

Temporal Preference Concepts and their Functions in a Large Language Model cites this paper.

Temporal Preference Concepts and their Functions in a Large Language Model Understanding (Un)Reliability of Steering Vectors in Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T17:03:44.315006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T17:03:44.315006Z digest=sha256:c770e4866033c8996d57fc5349af9d4799219ad6eb5356474e8cc4b167d43fca

Observation d460774b-6a78-4912-8e56-f903b96b75df · inbound

Adversarial Robustness of Activation Steering in Large Language Models cites this paper.

Adversarial Robustness of Activation Steering in Large Language Models Understanding (Un)Reliability of Steering Vectors in Language Models

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:37:09.432517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T22:30:36.839964Z digest=sha256:5618e24a3cbf3987d05994cb1f119ac7d94eccd99148e896df5f87c7e8227d35

Observation 4d6c0238-a48e-4825-ae00-cabcc32cf3ed · inbound

Detecting and Controlling Sycophancy with Cascading Linear Features cites this paper.

Detecting and Controlling Sycophancy with Cascading Linear Features Understanding (Un)Reliability of Steering Vectors in Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:49:57.005467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T01:24:56.600219Z digest=sha256:995d70d7c7efde7a46313c5f42d59fb5ed7857c8d62be6de99911cefdbe837ca

Observation ab15ab25-369b-452d-91c7-3fb5f45bc50c · inbound

Conditional Optimal Bridge for Riemannian Activation Steering cites this paper.

Conditional Optimal Bridge for Riemannian Activation Steering Understanding (Un)Reliability of Steering Vectors in Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-14T11:08:17.571722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T11:08:17.571722Z digest=sha256:98f85fb3d260345bb0e3b1decb8ac8ac70adc4122a2b28eeafc8513943294599