Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 38 inbound Pith citation observations for arXiv:2501.17148.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T00:27:20.702277Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 7a9fc4e6-95b4-4fb6-a4ba-502bb7e25f20 · inbound
Denoising Concept Vectors with Sparse Autoencoders for Improved Language Model Steering AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31cea893-d830-4649-a4bb-38ba2a5f3b78 · inbound
Steering Large Language Models for Machine Translation Personalization AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f67b55c0-fde2-4415-afdd-9906c7cd1c92 · inbound
Evaluating Steering Techniques using Human Similarity Judgments AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bbd5b28-16ef-422c-85fd-6ec1fde77f96 · inbound
Improved Representation Steering for Language Models AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23e4f887-4e7a-45ad-ad63-9c1688b592c9 · inbound
Aligned but Blind: Alignment Increases Implicit Bias by Reducing Awareness of Race AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a44a648d-3daf-4ca6-80dc-25d7bcf6f345 · inbound
Linear Representation Transferability Hypothesis: Leveraging Small Models to Steer Large Models AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de4cbd8d-1c7a-408d-9bbd-93ecc67f2526 · inbound
HyperSteer: Activation Steering at Scale with Hypernetworks AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f410453-9292-436c-a201-524a7700f897 · inbound
Fine-Grained Interpretation of Political Opinions in Large Language Models AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48448f35-3dc6-4cb3-b0e2-fd117fefe50a · inbound
Resa: Transparent Reasoning Models via SAEs AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa3cfd68-4461-42d7-9c7b-067819b1df5e · inbound
From Emergence to Control: Probing and Modulating Self-Reflection in Language Models AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 686e0ed0-7129-44b5-add7-9373dae8b45e · inbound
Position: Use Sparse Autoencoders to Discover Unknowns AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecec1d06-3b49-41fc-859c-0a010f994c08 · inbound
Insights into a radiology-specialised multimodal large language model with sparse autoencoders AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f90aaef-7103-4c7f-a616-dc6de6694890 · inbound
Sparse but Wrong: Incorrect L0 Leads to Incorrect Features in Sparse Autoencoders AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d6046e9-9891-47b7-b248-19a4ac46b9a4 · inbound
Sealing The Backdoor: Unlearning Adversarial Text Triggers In Diffusion Models Using Knowledge Distillation AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64bae72e-6f2d-45e3-a86e-bcd48d11aca8 · inbound
Investigating Symbolic Triggers of Hallucination in Gemma Models Across HaluEval and TruthfulQA AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fef9a84f-9bd2-402d-979e-78f556efb848 · inbound
VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a29067e-97b4-4317-a117-068745e3b2cc · inbound
The SuperActivator Mechanism: Transformers Concentrate Reliable Concept Signals in the Tail AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25cc1439-9855-45b6-90e2-52c1d205dc7e · inbound
Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 999f74ea-7573-496a-b670-863bce46d67f · inbound
When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c9542428-9d16-464c-9340-0a67098a26b5 · inbound
When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cbcdd5f3-9ca9-4d3e-921f-0895e37a543a · inbound
When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dc19f809-2c50-4736-a7d6-263c18fb155f · inbound
When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cf6380e4-61c1-493d-9226-07ebcd27f71b · inbound
The Open-Box Fallacy: Why AI Deployment Needs a Calibrated Verification Regime AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 03000c2e-bb75-415a-8a01-5338990b0c2b · inbound
Stories in Space: In-Context Learning Trajectories in Conceptual Belief Space AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 109
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5273eeed-d53c-4b66-aa4e-adcbe3a4aacb · inbound
WriteSAE: Sparse Autoencoders for Recurrent State AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d5d1b455-8957-4939-afef-686d00f5552e · inbound
When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 420f821d-c186-40f2-8e77-c80e1bd8075a · inbound
When Is Rank-1 Steering Cheap? Geometry, Granularity, and Budgeted Search AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 037a966f-e60b-465f-aa09-3c95ed64ef6f · inbound
Are Sparse Autoencoder Benchmarks Reliable? AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f63048b1-1502-4a22-91ff-0f9842ba83ec · inbound
Activation Steering for Synthetic Data Generation: The Role of Diversity in Downstream Safety Detection AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e55ea6d-c6f2-4dd1-a1a5-9344dddd060a · inbound
Temporal Preference Concepts and their Functions in a Large Language Model AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 117
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1b40c27b-9c83-451e-bbbd-1341a9ee321f · inbound
Temporal Preference Concepts and their Functions in a Large Language Model AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 117
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3ba96c8-ca5b-4522-8973-6c0e25c6d87e · inbound
When is Your LLM Steerable? AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e79bbe08-8e29-4e61-8e50-6f9acde5ce5b · inbound
Anatomy of Post-Training: Using Interpretability to Characterize Data and Shape the Learning Signal AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 105
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 71cc1d01-e169-4ee8-bfb5-708a7379f935 · inbound
Size Doesn't Matter: Cosine-Scored Sparse Autoencoders AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 67b37175-3b56-4910-984d-8f416ca019f6 · inbound
Retrieval is Enough: Training-Free Interpretability with a Tool-Using Agent AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 608fb2c6-208f-4036-b5af-45253a520b6a · inbound
Probabilistic Concept-Aware Steering for Trustworthy LLM Inference AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2b9c053-81bb-4ef3-8471-255f8b679817 · inbound
ODRA: Synthesizing Cognitive Behavioral Therapy Sessions with Structured Chain-Of-Thought and Dynamic Patient Resistance AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52a1fa14-295f-46b4-87ae-5d6cc4258b0b · inbound
CircuitSteer: Geometrically Aligned Multi-Layer Steering via Sparse Autoencoder Circuits AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.