Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 44 inbound Pith citation observations for arXiv:2311.03348.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T00:12:04.882237Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T19:50:11.090116Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 048aaf97-31c3-4526-bfc8-a3773f238a7e · inbound
Jailbreaking Black Box Large Language Models in Twenty Queries Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 84096182-e022-4442-8f53-f1c2d080ab13 · inbound
Dr. Jekyll and Mr. Hyde: Two Faces of LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 49b0f1a6-1594-49cc-b309-7617cf3e94c1 · inbound
A StrongREJECT for Empty Jailbreaks Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dc4002bc-dea2-4b34-b549-7dd66ece4c54 · inbound
JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation db505cbd-6cd1-4674-8b49-730762ea7099 · inbound
Jailbreak Attacks and Defenses Against Large Language Models: A Survey Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 941745b9-75cd-4f8c-9a7d-aa987e2725d1 · inbound
Jailbreaking LLMs' Safeguard with Universal Magic Words for Text Embedding Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d848e5d-3895-4631-9e7a-b58f4573cde8 · inbound
Agents Are All You Need for LLM Unlearning Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 836769b0-789b-45a2-aa2e-1c59c43e35f4 · inbound
Jailbreaking with Universal Multi-Prompts Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc1c11dd-ec9f-45ee-8622-dd26f0a580b0 · inbound
Position: Adversarial ML for LLMs Is Not Making Any Progress Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76c2f389-29c0-4cc6-9ff1-f7d42a703277 · inbound
QueryAttack: Jailbreaking Aligned Large Language Models Using Structured Non-natural Query Language Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4c1861b-9875-4628-86e9-e6c8780e54de · inbound
Lifelong Safety Alignment for Language Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 995c9188-0317-4fba-a6a0-834c209a5684 · inbound
Jailbreak Distillation: Renewable Safety Benchmarking Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c0bbdf6-817c-4bd0-a666-601b851f6a74 · inbound
Adversarial Attacks on Robotic Vision Language Action Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c351043-6e22-4ba2-8063-e415567b38df · inbound
The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4caf03b7-64e8-467e-aa4e-56422e420bb6 · inbound
SecurityLingua: Efficient Defense of LLM Jailbreak Attacks via Security-Aware Prompt Compression Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d906230-1a4e-4b3f-b9a8-854bff233fbf · inbound
SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5942f2a1-cda5-4fab-8a27-6ec6c3c2b392 · inbound
VERA: Variational Inference Framework for Jailbreaking Large Language Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74e33e9a-f10b-4f6d-88cb-9af4f587b0f0 · inbound
Linearly Decoding Refused Knowledge in Aligned Language Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b853babb-024f-485f-bda8-0d6e9d553ba5 · inbound
PRM-Free Security Alignment of Large Models via Red Teaming and Adversarial Training Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02860e26-ed0f-46bf-9da8-bf95f39ad648 · inbound
From Seed to Harvest: Augmenting Human Creativity with AI for Red-teaming Text-to-Image Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bb6def3-2f52-4d35-8150-66280a73bd78 · inbound
MOCHA: Are Code Language Models Robust Against Multi-Turn Malicious Coding Prompts? Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0049bed-32c9-4638-81e1-7145f1c03b4c · inbound
The Emotional Baby Is Truly Deadly: Does your Multimodal Large Reasoning Model Have Emotional Flattery towards Humans? Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a625e3c-4afa-4f3b-b1e3-eb81901693df · inbound
Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25ebd5fc-712f-4be4-8ac9-6177452bcd62 · inbound
On Surjectivity of Neural Networks: Can you elicit any behavior from your model? Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51bae0dc-effd-4724-9804-6e3ad33b4bdb · inbound
GUARD: Guideline Upholding Test through Adaptive Role-play and Jailbreak Diagnostics for LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e652e991-c4d7-451c-a0bf-7f831e253e4a · inbound
Probabilistic Modeling of Latent Agentic Substructures in Deep Neural Networks Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6fcd0746-d538-4497-a298-36eeebcd2e89 · inbound
GAMBIT: A Gamified Jailbreak Framework for Multimodal Large Language Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b1f54b27-8df9-45c9-bd09-5e105446ab8f · inbound
Stop Tracking Me! Proactive Defense Against Attribute Inference Attack in LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8cb8c199-c5ba-41cb-8c8e-6b0e44334727 · inbound
State-Dependent Safety Failures in Multi-Turn Language Model Interaction Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56730341-2f6b-4825-99bc-171d507ae8ff · inbound
Conflicts Make Large Reasoning Models Vulnerable to Attacks Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7f80a41d-4886-426a-9dd5-1acd3c6d89a2 · inbound
Too Nice to Tell the Truth: Quantifying Agreeableness-Driven Sycophancy in Role-Playing Language Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 92b5cdac-c2c9-4c7f-a753-0cf32915931b · inbound
VoxSafeBench: Not Just What Is Said, but Who, How, and Where Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0cf46a83-bcf0-4960-99a5-9c5e20c551b4 · inbound
Disentangling Intent from Role: Adversarial Self-Play for Persona-Invariant Safety Alignment Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 429303f9-8b9c-40b1-bb68-8dd7a4bfb69a · inbound
On the Hardness of Junking LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 02566d56-245c-4f38-abd7-e26852bcab41 · inbound
ContextEcho: A Benchmark for Persona Drift in Long Agentic-Coding Sessions Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0fe7bbb2-338d-4b87-bf0e-a68e330e07ff · inbound
Quality-Diversity Evolution for Discovering Diverse Vulnerabilities in LLM Safety Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ad43cf6e-9b85-4c0a-8d00-384d15da4c5e · inbound
CHASE: Adversarial Red-Blue Teaming for Improving LLM Safety using Reinforcement Learning Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7e0f354e-5d7b-471a-b39d-e07f8a5b6c28 · inbound
SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ee858e8a-052b-4b27-921c-038480946b72 · inbound
A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3f5ce316-7ac9-4e84-bc73-956f6b4b9924 · inbound
A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54147477-564f-43bc-b241-f968bbf1548c · inbound
Cognitive Firewall: A Proactive, Zero-Trust, Multi-Gate Framework for LLM Safety Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 10a14e1f-2063-4e44-93e4-a61675755567 · inbound
A Scalable Approach to Evaluating Moral Sensitivity in LLMs Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 157
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ae2cb66-26b5-46ae-99e1-d23cfab18d5b · inbound
Execution-Grounded Security Testing for Coding Agents in Software Engineering Pipelines Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39690abe-ff05-411e-9ae9-0698e268e704 · inbound
Role Steering of Language Models for Social Simulations Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.