Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T01:03:22.328312Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 6 inbound Pith citation observations for arXiv:2506.12152.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T01:03:22.328312Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T03:50:15.977348Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T00:59:21.346717Z
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 91d74f30-bc4c-459d-9d54-e29fb3303496 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Abdul, J
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7309861d-5495-4868-847b-d9a5f7b77729 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Studying Large Language Model Generalization with Influence Functions
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5660c1af-c3fc-4483-a960-a1a2a1234d4b · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 065723a3-5382-48fc-8277-e83830fc674f · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99aaa16c-8ae6-4aa2-b35e-a55530ecda6a · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability doi: 10.1145/3564240
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b20b3980-7f26-480f-9757-23a47c91446e · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability An Approach to Technical AGI Safety and Security
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5584384c-eafb-41cf-9f8c-642c945093dd · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Open Problems in Mechanistic Interpretability
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2e9c52c-864f-43da-97e9-fc5dcca42dd8 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7bdb23f-159c-4c85-a2eb-c10b77497794 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability doi: 10.18653/v1/D16-1159
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb7f58e0-4da1-417f-8b70-2b500455401a · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41a21ca5-2781-41f9-a691-92e990be6664 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability A Roadmap to Pluralistic Alignment
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9a5a264-c583-455f-8512-4cb5075a344e · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9a5d7c8-0322-4597-9f3f-d296e7bb0d50 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Large Language Models Are Human-Level Prompt Engineers
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9c4d9f5-7700-4a11-892a-e4fe517619da · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability URLhttps://www.ncbi
Reference 1974
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28924536-ca9b-4f84-a1e7-c13635f86df4 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Merriam-webster, Accessed on 2025-05-22
Reference 1982
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a2856cf9-3960-4f23-8784-086a987b8959 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Designing a Dashboard for Transparency and Control of Conversational AI
Reference 1993
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecd4a1a0-0a54-48aa-9cdd-44aae1d59611 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Alignment faking in large language models
Reference 2012
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdaa9f40-15d2-4727-9204-0c22b5514aad · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Constitutional AI: Harmlessness from AI Feedback
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 402b051c-2f8b-494d-9fe3-1b581854a724 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability doi: 10.18653/v1/W16-2524
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f079495-60a1-40d6-a715-75f588fa91e0 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Auditing language models for hidden objectives
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c08c3d79-8d55-437c-8cf4-6c9365c936ab · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Sampling Method for Fast Training of Support Vector Data Description
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 56ab1b47-fcb6-4cfd-99c8-6d4a8185974d · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Unresolved cited work
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8f150441-741f-4da9-b5ca-13a1248d9ad8 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability AutoPrompt: Eliciting Knowledge from Language Models with Automatically Generated Prompts
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6adca01-aeda-4876-9a8b-2b6856c8ceec · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Unresolved cited work
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c147767b-17da-4232-8d06-db94f7cb54d5 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Unresolved cited work
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 009c88f1-a1f1-46ef-8ef4-504a046f3b1c · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Progress measures for grokking via mechanistic interpretability
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a287e320-3fd1-421a-b9ef-8b7bf845bac5 · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Challenges in Human-Agent Communication
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bd2d3f5-d96d-47ae-85b0-ee97be0b6d7f · outbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability We Can't Understand AI Using our Existing Vocabulary
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8949abfc-dd1c-4b41-85b6-5aed53fdaacc · inbound
Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs Because we have LLMs, we Can and Should Pursue Agentic Interpretability
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 715bf029-5944-4d4b-a090-5a2e8a47486a · inbound
From Features to Actions: Explainability in Traditional and Agentic AI Systems Because we have LLMs, we Can and Should Pursue Agentic Interpretability
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9b42aa7-1170-456f-92e8-9d6809daa34b · inbound
A Geometric View for Understanding Concept Learning and Neuron Interpretation in Sparse Autoencoders Because we have LLMs, we Can and Should Pursue Agentic Interpretability
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d0440bd1-9207-4522-b016-ff15cda7a584 · inbound
Uncertainty Decomposition for Clarification Seeking in LLM Agents Because we have LLMs, we Can and Should Pursue Agentic Interpretability
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6e828df5-a319-45ff-a816-957d22a9a462 · inbound
The Curse of Multiple Mediators: Hidden Interaction Effects in Activation Patching Because we have LLMs, we Can and Should Pursue Agentic Interpretability
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 78706c2d-689e-48bf-a2ca-7c96dc239c67 · inbound
Reality Monitoring in Large Language Models: Self-Knowledge That Transforms with Conversation Memory Because we have LLMs, we Can and Should Pursue Agentic Interpretability
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.