Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2311.05553.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:58:32.538591Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T05:39:40.653733Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation d3a1fce6-3ae4-4331-beba-8a22e0d1e594 · inbound
A StrongREJECT for Empty Jailbreaks Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d438dfad-e3f6-46b6-b417-7baa05af7aa5 · inbound
LLM Agents can Autonomously Exploit One-day Vulnerabilities Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 08b13389-b693-485c-9a26-527420eaa3d1 · inbound
Refusal in Language Models Is Mediated by a Single Direction Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 205
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d76278a6-65bb-4712-b4da-f9c535c79cde · inbound
Jailbreak Attacks and Defenses Against Large Language Models: A Survey Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 111
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation da4add40-2fa4-4b49-8281-4abcaf742df1 · inbound
Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 177
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 382b390f-a3a8-4031-8574-31eb66cb9f27 · inbound
Vulnerability-Aware Alignment: Mitigating Uneven Forgetting in Harmful Fine-Tuning Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab1d7d35-a32d-4fb2-9219-90097f0a8d86 · inbound
SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2fffb31-5954-4016-ac38-f43310662a1b · inbound
Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d59c73aa-0579-466e-b148-13041ea73cf9 · inbound
S3LoRA: Safe Spectral Sharpness-Guided Pruning in Adaptation of Agent Planner Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cee53ca1-667e-4e37-bb0f-bb69aa9df89f · inbound
Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d1f1798-edec-42e1-96a8-a70f3a11a1e9 · inbound
Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60840032-8432-4455-8d31-8136cb96982e · inbound
MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4d2b640-0807-482c-98e5-f9235fcac7af · inbound
Paladin: Defending LLM-enabled Phishing Emails with a New Trigger-Tag Paradigm Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f08c312-b145-4507-a47f-db9af96708e9 · inbound
Robust Policy Optimization to Prevent Catastrophic Forgetting Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c9d07f83-bfc7-44a4-aacb-e4d1a4a6dac3 · inbound
The Possibility of Artificial Intelligence Becoming a Subject and the Alignment Problem Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 192a1bc1-ec29-49b7-bb81-ab010516d622 · inbound
The Possibility of Artificial Intelligence Becoming a Subject and the Alignment Problem Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7e7202d-0fa1-44cd-a743-f8d94550050b · inbound
Representation-Guided Parameter-Efficient LLM Unlearning Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 023e10d7-9b08-4f5e-b37d-f7027db4232a · inbound
You Snooze, You Lose: Automatic Safety Alignment Restoration through Neural Weight Translation Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 146
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1cf88bd4-e26a-4ec8-ae27-9674b19c91ac · inbound
Open Weight AI Models Require Proportional Evaluation Approaches Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0bc464dc-398e-45a6-a3c0-388d06036cc7 · inbound
How LLM Task-Adaptation Reshapes Alignment: A Multi-dimensional Study of Behavioral and Representational Drift Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c0ce0e2-6c99-4603-94ce-9450a6693d79 · inbound
AI Security Priorities: A Field-Wide Agenda Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.