Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:37:08.568150Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2506.10236.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:37:08.568150Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6224a955-ca34-442b-aeb5-51011391d676 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Who's Harry Potter? Approximate Unlearning in LLMs
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 784e5666-a270-4ef5-a7d0-1fb1b4a0cedc · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods 6 Ryan Greenblatt, Fabien Roger, Dmitrii Krasheninnikov, and David Krueger
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 063ad588-eda8-45ea-9ed0-4369c984792e · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Mistral 7B
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa825a17-919c-4dcf-a4ee-b15251b88615 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Eight Methods to Evaluate Robust Unlearning in LLMs
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c4baf1d-82b2-43f2-ad2d-084173759af0 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods TOFU: A Task of Fictitious Unlearning for LLMs
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81467daf-364b-48ba-afc0-c08fe4e21be5 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods McKinney, Anvith Thudi, Juhan Bae, Tara Rezaei Kheirkhah, Nicolas Papernot, Sheila A
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3f40819a-a57e-4375-ae14-5ec259a731e8 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Can Sensitive Information Be Deleted From LLMs? Objectives for Defending Against Extraction Attacks
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21c52f53-8e11-467b-b0bb-db5c958e57cc · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods tinyBenchmarks: evaluating LLMs with fewer examples
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e31cf72-4903-4ae0-b01f-ba26f6d10686 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Latent Adversarial Training Improves Robustness to Persistent Harmful Behaviors in LLMs
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f7bb65d-affd-40f1-97b4-2e08bda272ab · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods UnUnlearning: Unlearning is not sufficient for content regulation in advanced generative AI
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72cbd017-c19c-46bb-ae8f-632a1a95f874 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Tamper-Resistant Safeguards for Open-Weight LLMs
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd246f8c-77b8-44ad-b04e-a47d6ebc2c5e · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods AI Sandbagging: Language Models can Strategically Underperform on Evaluations
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d61c81e3-d491-44ad-b92a-caff49b54915 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Jailbroken: How Does LLM Safety Training Fail?
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2497b4e0-902a-4fe1-87aa-e909ac647df9 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods In-Context Learning Can Re-learn Forbidden Tasks
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e7b22fde-4c56-45b1-8e2e-19c8d5f09556 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Low-Resource Languages Jailbreak GPT-4
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9b446f5-9593-4ab4-8bb8-172c1d9c5ae5 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Towards Robust Knowledge Unlearning: An Adversarial Framework for Assessing and Improving Unlearning Robustness in Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation daf4a539-e064-49db-90ab-d9a500267018 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Improving Alignment and Robustness with Circuit Breakers
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff311a48-1010-47e9-b739-5526929b2600 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods [2024], a subset of 100 data points selected from MMLU (Massive Multitask Language Understanding) Hendrycks et al
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f746d274-9a40-4499-b45c-c5bba14a7e2d · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Right Format
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 381b5830-0b54-4e86-b56f-edf92e64f8bd · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods The Elicitation Game: Evaluating Capability Elicitation Techniques
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 546dad7a-a2f3-4619-85ee-905b22e762e5 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Continual Learning and Private Unlearning
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e0fc9f6-87cb-4b06-91ca-970599e27cbd · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Erasing Conceptual Knowledge from Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ac562e3-c5b1-46c0-8e04-7102890168a9 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Model Tampering Attacks Enable More Rigorous Evaluations of LLM Capabilities
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c1bd043-7b6f-478f-bcbb-b698ff529c13 · outbound
Prompt Attacks Reveal Superficial Knowledge Removal in Unlearning Methods Does Unlearning Truly Unlearn? A Black Box Evaluation of LLM Unlearning Methods
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.