Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:01:11.473485Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 8 inbound Pith citation observations for arXiv:2506.00782.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:01:11.473485Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T16:21:26.984437Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T04:16:34.709256Z
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9a1e72f4-36bd-4f13-ba4c-6fb71b2c8699 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Claude-3.5-sonnet, 2024
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a57cad7f-0a66-442d-ba67-07d698debdec · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Diverse and Effective Red Teaming with Auto-generated Rewards and Multi-step Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c83958e-51f1-4715-841a-38e611cdb273 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Bhardwaj, D
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bd7a3d64-e8c8-44d7-8978-302a0a69abae · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning On the Opportunities and Risks of Foundation Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f393ed59-e8d9-45e9-8660-c820a41aa1e5 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Jailbreaking Black Box Large Language Models in Twenty Queries
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b364c890-f672-4494-8c4e-0c3ca41f3b99 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Chiang, L
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3a3d76af-8b39-4115-a880-59665351a6f2 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 673ba450-cba6-4eeb-b443-628662efb98e · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c18d2fbe-eeb3-4382-a8db-52aef4f2b1d5 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning The Llama 3 Herd of Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d985b050-1d04-41e4-95bc-c5fb7d70cc38 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 072fc78b-14de-47d8-9eb1-8f1452568d74 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ca9c6f7a-6ddd-4ef9-9e2c-e833b8540891 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Best-of-N Jailbreaking
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eefbd577-524a-4cc1-a8d0-3faa2c8147af · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f96588aa-b5a8-47c9-a3a3-77cb239334db · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Jiang, K
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8b7839bf-1144-4e59-9590-e49c9f2bbb38 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Learning diverse attacks on large language models for robust red-teaming and safety tuning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9d7b985-ec8f-438f-b860-4a5163b078c8 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation caa11580-6c71-4128-ad4f-253b6c47ab5d · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AmpleGCG: Learning a Universal and Transferable Generative Model of Adversarial Suffixes for Jailbreaking Both Open and Closed LLMs
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33b04a5e-5201-4103-a505-f126f201da8a · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AutoDAN-Turbo: A Lifelong Agent for Strategy Self-Exploration to Jailbreak LLMs
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79503cc3-a8ae-4757-97cc-c009a116ff06 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7ef22d1b-236d-4d76-b9a5-70094e50e563 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Auto-RT: Automatic Jailbreak Strategy Exploration for Red-Teaming Large Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f76d43c-b567-41d5-8d90-ffe9b0a7e40a · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Mazeika, L
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b1e15a24-cb77-4191-9ab0-92c61ff84c58 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Mehrotra, M
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 566053bd-80f3-4d37-853f-df3a49293b28 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Gpt-3.5 turbo, 2023
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 626e1ff1-468b-4aaa-bfb7-e71a93d30c8f · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Gpt-4o system card, 2024a
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 08782bc4-4c05-447a-8f7a-e6054cf6024a · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8db39e9-155c-47b6-9fc9-841594d08b36 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Perez, S
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8608203b-ab0d-4b7d-9e6d-5f5952b7c7aa · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning ToolRL: Reward is All Tool Learning Needs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a61ec825-55bf-40eb-92e3-e9999b760f4e · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Samvelyan, S
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2fe3d659-1d9e-4572-83d3-d16129c3df13 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 337364ca-5d04-4c9d-aa9e-140249eaeb30 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Shaikh, H
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 33a0d048-074e-4b90-8607-1d344d8b45c1 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1da0f343-f8de-4521-ae3e-d911e0b7fecd · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c80dcfb9-4c3d-4a0d-b3d0-46c2a6f5d20a · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20bb5f9f-9282-4eea-9926-6fe1ae134672 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Wang and K
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9116b090-aef8-4210-95f5-bf89a5cdba32 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Qwen2 Technical Report
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 964183fa-aa52-49f2-927f-c5173047915a · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d46e43ee-3901-486a-9347-9043cb9922c4 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning STAIR: Improving Safety Alignment with Introspective Reasoning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9be8f1ce-bb31-4409-9aea-d7454a72728f · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0aeb746e-80aa-4229-83f4-d714ffbf2ddf · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Toward Optimal LLM Alignments Using Two-Player Games
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 362e0e77-debc-4ddb-a8c9-f1c8227e1e3d · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AutoRedTeamer: Autonomous Red Teaming with Lifelong Attack Integration
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed3ab69c-a3f5-45fb-a68a-4206b040c247 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Purple-teaming LLMs with Adversarial Defender Training
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc43c2c8-46e6-48aa-b081-3da843160ed5 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d3ff433-976c-4255-ac9a-da42342f056e · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning After obtaining the filtered 2k samples, we prompt the Qwen2.5-7B-Instruct model to imitate the sample attack as an example
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e1cf1b40-4758-4b2e-8d33-9e64c96a27d7 · outbound
Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3caa8e0b-18b1-4a1e-a875-266a80aa3761 · inbound
Stable-GFlowNet: Toward Diverse and Robust LLM Red-Teaming via Contrastive Trajectory Balance Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4426f7f5-4c93-4191-8186-f553eb93f59c · inbound
Internalizing Safety Understanding in Large Reasoning Models via Verification Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4645865c-d491-420f-917f-eb479a6b126e · inbound
Self-ReSET: Learning to Self-Recover from Unsafe Reasoning Trajectories Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a31831b8-9265-4c4e-9e19-3609eaaf51d2 · inbound
Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e777a612-105b-4bf5-bd5a-b1cd2db8b8f5 · inbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b78426c-7ced-42f2-b964-4f713e454373 · inbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8ab7de0-ef03-447a-9f60-7378374f6c54 · inbound
MJ: Multi-turn LLM Jailbreaking via Decomposed Credit Assignment Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b3f8d1d-b25c-4b2c-a058-d53d4c66aeff · inbound
An Early Warning of Emerging Biosecurity Risks in Frontier LLMs Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.