Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2306.09442.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:35:20.649185Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
7
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation e6f0ac5f-7df8-4bf4-a62d-954b350d54a1 · inbound
A Comprehensive Overview of Large Language Models Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 179
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 21b70197-ee2c-45d3-97f5-10495b97384d · inbound
Baseline Defenses for Adversarial Attacks Against Aligned Language Models Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7f4390b7-62ee-445a-8216-340d244f8201 · inbound
Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d6ab93bd-c519-45f4-a798-b834acb4f06f · inbound
Agent AI: Surveying the Horizons of Multimodal Interaction Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation febf8415-c6bf-42ff-b85c-77200f92c9da · inbound
TrustLLM: Trustworthiness in Large Language Models Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 231
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3eb0970a-203c-4127-8961-9e16f4df2c85 · inbound
Jailbreak Attacks and Defenses Against Large Language Models: A Survey Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9b32b33d-7756-4ec4-913e-7c6d1be463bd · inbound
Adversarial Preference Learning for Robust LLM Alignment Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae34b8d3-fbac-43ec-a207-7dba2ef5b36d · inbound
RedRFT: A Light-Weight Benchmark for Reinforcement Fine-Tuning-Based Red Teaming Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfba1cf5-35fd-49e4-81d5-182a7d296090 · inbound
FORTRESS: Frontier Risk Evaluation for National Security and Public Safety Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1c03a4e-92e4-4a51-9c7c-61118fc31784 · inbound
Kaleidoscopic Teaming in Multi Agent Simulations Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54a66195-b915-43e4-9959-84fd95786d5c · inbound
SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee47c9ff-dcb0-43a9-a097-9fad4e0e75d8 · inbound
PRM-Free Security Alignment of Large Models via Red Teaming and Adversarial Training Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fa9f4d3-d632-40d9-a50d-d4ed6826d947 · inbound
Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 532f346b-bf62-476d-9d95-24477cb4d118 · inbound
Quantized but Deceptive? A Multi-Dimensional Truthfulness Evaluation of Quantized LLMs Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb7228e8-a91f-4b61-ae2a-d1c55a42c3b9 · inbound
Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f56caf1-de30-425d-a0ec-e7d7fabde360 · inbound
MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cf155fc-6461-4c27-a3bc-1c990bfdb00c · inbound
Tailored untruths: How personalisation challenges LLM safeguards Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab94034e-cb58-4d6f-acc7-7f9ffdad0f68 · inbound
Learning Uncertainty from Sequential Internal Dispersion in Large Language Models Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d8955db2-bcdd-499c-b8ec-3b9aad7f8b3e · inbound
Reverse Constitutional AI: A Framework for Controllable Toxic Data Generation via Probability-Clamped RLAIF Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 27e374e1-026f-4884-9734-44dca0ac509e · inbound
A Systematic Investigation of RL-Jailbreaking in LLMs Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c8e15fe-8f07-42ae-8d0c-98f83597482b · inbound
A Systematic Investigation of RL-Jailbreaking in LLMs Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 013aef7e-a486-4a92-b0c6-1e03faf6af9e · inbound
A Systematic Investigation of RL-Jailbreaking in LLMs Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b326c0f5-09b8-470d-bb57-d1c0b650a771 · inbound
Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 207
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d7bd52b0-f8a0-451f-9ea4-8ab3a91f5838 · inbound
Formal Methods Meet LLMs: Auditing, Monitoring, and Intervention for Compliance of Advanced AI Systems Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 622045b0-0c14-4d76-a948-691036c06a9d · inbound
Boosting Self-Consistency with Ranking Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 136
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c911afea-adfd-4ed9-aa69-f2d0abe6abf6 · inbound
Data Selection Through Iterative Self-Filtering for Vision-Language Settings Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c5b64cb-10fe-4708-8cb7-f0e89a5b8e2a · inbound
On-Policy Self-Distillation with Sampled Demonstrations Reduces Output Diversity Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c257477e-7e72-4329-a983-e8ba226c1af0 · inbound
Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40e14c18-6f9e-4d12-aa77-8bfc6d0d6d39 · inbound
Just Keep Prompting: Evaluating Repetitive Socratic Prompting in VLMs Explore, Establish, Exploit: Red Teaming Language Models from Scratch
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.