Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 25 inbound Pith citation observations for arXiv:2411.11296.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:03:01.923666Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 8971053d-2e39-4d5b-9596-5976a58bb55b · inbound
Position: Mechanistic Interpretability Should Prioritize Feature Consistency in SAEs Steering Language Model Refusal with Sparse Autoencoders
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db905e8f-0f2c-49fb-9f55-f33b8a67a59d · inbound
Incorporating Hierarchical Semantics in Sparse Autoencoder Architectures Steering Language Model Refusal with Sparse Autoencoders
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbeb1e1e-31e7-498f-af5f-ca4aa00efc3b · inbound
Interpretation Meets Safety: A Survey on Interpretation Methods and Tools for Improving LLM Safety Steering Language Model Refusal with Sparse Autoencoders
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 079557b3-ec61-4220-b289-83d8193b0886 · inbound
Resa: Transparent Reasoning Models via SAEs Steering Language Model Refusal with Sparse Autoencoders
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48c18206-9944-4304-8f43-0c8bd96911c9 · inbound
Insights into a radiology-specialised multimodal large language model with sparse autoencoders Steering Language Model Refusal with Sparse Autoencoders
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 043bb4b5-df6b-48dc-923e-a044138069e9 · inbound
Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM Steering Language Model Refusal with Sparse Autoencoders
Reference 231
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77d84fa0-46e6-4bbc-826a-a011a1d5b69e · inbound
Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Steering Language Model Refusal with Sparse Autoencoders
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce18123b-8d33-4f35-90a6-151542840ddf · inbound
SATORI: Static Test Oracle Generation for REST APIs Steering Language Model Refusal with Sparse Autoencoders
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf2db5bc-92c8-4a3e-aba5-1c89c5bf2c6d · inbound
Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection Steering Language Model Refusal with Sparse Autoencoders
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc858d3f-411b-473a-9127-b5651c15cc71 · inbound
Beyond I'm Sorry, I Can't: Dissecting Large Language Model Refusal Steering Language Model Refusal with Sparse Autoencoders
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f5868bf6-2a9a-49a2-8707-729707477d46 · inbound
The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook Steering Language Model Refusal with Sparse Autoencoders
Reference 155
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9060515b-c19f-4604-a3ca-b523abf2b4f4 · inbound
Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs Steering Language Model Refusal with Sparse Autoencoders
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bb0164bb-25f3-4847-a5b2-40670026e08e · inbound
Towards Understanding the Robustness of Sparse Autoencoders Steering Language Model Refusal with Sparse Autoencoders
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation adfbfdad-b81c-4b55-b90d-9162cfa2bd44 · inbound
Estimating Tail Risks in Language Model Output Distributions Steering Language Model Refusal with Sparse Autoencoders
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3ca12e34-41b2-461b-ba5e-91aa0972adeb · inbound
Don't Lose Focus: Activation Steering via Key-Orthogonal Projections Steering Language Model Refusal with Sparse Autoencoders
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3a5c6831-ecd7-40a7-a79a-4013d95f35ef · inbound
LLM Advertisement based on Neuron Auctions Steering Language Model Refusal with Sparse Autoencoders
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a00562e3-4f2b-4052-b15e-f2f31e683351 · inbound
When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search Steering Language Model Refusal with Sparse Autoencoders
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5ea39289-0c37-4c0c-8513-f81a05d3012e · inbound
When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search Steering Language Model Refusal with Sparse Autoencoders
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ec170ce0-4d8b-4f1b-8070-d72265cd1409 · inbound
Ablating Safety: Mechanisms for Removing Alignment in Language Models for Security Applications Steering Language Model Refusal with Sparse Autoencoders
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b9b04e5e-0d17-4029-be26-b96cc53970ed · inbound
Steered Generation via Gradient-Based Optimization on Sparse Query Features Steering Language Model Refusal with Sparse Autoencoders
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation afcfa71d-630e-48f5-afcc-def0ff4f85ff · inbound
Palette: A Modular, Controllable, and Efficient Framework for On-demand Authorized Safety Alignment Relaxation in LLMs Steering Language Model Refusal with Sparse Autoencoders
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b0dcb2c8-50b8-4dca-b7f6-0fad16a9eb44 · inbound
Perplexity Can Miss SAE Feature Damage Under Quantization Steering Language Model Refusal with Sparse Autoencoders
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2a37f71a-0e2a-40e3-9b29-4b3a68983bb5 · inbound
Pre-Intervention Prediction of Sparse Autoencoder Steering Side Effects Steering Language Model Refusal with Sparse Autoencoders
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation de07c086-5ed4-4d00-9a96-fb9b97482d67 · inbound
At the Edge of Understanding: Sparse Autoencoders Trace The Limits of Transformer Generalization Steering Language Model Refusal with Sparse Autoencoders
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c1c16bb2-90b4-42f0-bfac-0c46ae57d38a · inbound
Forecasting With LLMs: Improved Generalization Through Feature Steering Steering Language Model Refusal with Sparse Autoencoders
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.