Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:50:25.717337Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 4 inbound Pith citation observations for arXiv:2507.02956.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:50:25.717337Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-29T01:28:07.221590Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T18:55:58.422476Z
23 of 23 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8c3456ac-1348-47f6-a910-eef0958bb212 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 476c410e-88e4-4910-9937-32e74c627486 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Well, that escalated quickly: The Single-Turn Crescendo Attack (STCA)
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8d36f44d-8bb5-4879-9bdf-42d0bc633329 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b4a4ebe-5b5e-4188-87ae-6913e3b6696e · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Constitutional AI: Harmlessness from AI Feedback
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cedc3e42-4251-41be-abd5-a43435295f02 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Jailbreaking Black Box Large Language Models in Twenty Queries
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c45739ee-df9e-4c02-9caf-85f8fda83311 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d78fe336-7b68-4781-a6de-0a99ff02065d · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Deliberative Alignment: Reasoning Enables Safer Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eba21bb8-2c2f-4933-94eb-e0b8a11402c2 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks COLD-Attack: Jailbreaking LLMs with Stealthiness and Controllability
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87a9aac9-c8c8-44dc-a605-a480afb89877 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Steering dialogue dynamics for robustness against multi-turn jailbreaking attacks, 2025
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6e2f740-a3b4-4596-b545-b38fd8c10bda · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks LLM Defenses Are Not Robust to Multi-Turn Human Jailbreaks Yet
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 218f2879-44f5-42f8-b73c-0f0c7aae61c5 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks X-boundary: Establishing exact safety boundary to shield llms from multi-turn jailbreaks without compromising usability, 2025
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca0ad785-49d1-417c-ab53-21df10b7f92d · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Tree of Attacks: Jailbreaking Black-Box LLMs Automatically
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a8ce7cb-43d4-49a9-9712-bc9a094941ff · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks PyRIT: A Framework for Security Risk Identification and Red Teaming in Generative AI System
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de570d9d-cb13-4aed-84ad-1348293e97ca · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Automated Red Teaming with GOAT: the Generative Offensive Agent Tester
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 986520c3-098d-47ef-84fd-10564fb86639 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Great, Now Write an Article About That: The Crescendo Multi-Turn LLM Jailbreak Attack
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d96455b2-84ed-4436-8021-783c7711c8d1 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Soft Prompt Threats: Attacking Safety Alignment and Unlearning in Open-Source LLMs through the Embedding Space
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b287634-43b3-4732-a2ec-965d9f8786e2 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Constitutional Classifiers: Defending against Universal Jailbreaks across Thousands of Hours of Red Teaming
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3aed3228-bacd-4e80-817f-1acc3b673ee6 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Taxonomy, opportunities, and challenges of representation engineering for large language models, 2025
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a125ad1-e518-4342-b435-8c35075d0f98 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Representation Bending for Large Language Model Safety
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fb5d3d7-f02d-4e4d-bae5-75e22877ec81 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Tempest: Autonomous Multi-Turn Jailbreaking of Large Language Models with Tree Search
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ec2eda3-b55d-4171-9724-fd207bff6ccd · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Representation Engineering: A Top-Down Approach to AI Transparency
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc2a659d-dbc5-46c1-9da7-b01d1cc11c60 · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b01ed218-98a2-470d-8d2b-fde3eb6d002b · outbound
A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks Improving Alignment and Robustness with Circuit Breakers
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15d0bef9-8dc8-4ab5-a86e-a1d3216ca85b · inbound
The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ad50daed-644a-472a-93fc-0ab8ed1ea288 · inbound
Mitigating Many-shot Jailbreak Attacks with One Single Demonstration A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bcb86915-e183-4ac1-a23d-d2f1285ca74b · inbound
From AI-Generated Content to Agentic Action: Security and Safety Threats in Generative AI A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 889f0505-845f-4f7d-b10c-838f7a34905f · inbound
On the Inseparability of Instructions and Data in Shared-Embedding Sequence Models A Representation Engineering Perspective on the Effectiveness of Multi-Turn Jailbreaks
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.