Pith. sign in

Paper Citation Record · LEDGER

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

As of 7 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 8 inbound Pith citation observations for arXiv:2506.00782.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00782 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:01:11.473485Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T16:21:26.984437Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T04:16:34.709256Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy13
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9a1e72f4-36bd-4f13-ba4c-6fb71b2c8699 · outbound

This paper cites Claude-3.5-sonnet, 2024.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Claude-3.5-sonnet, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:15.888902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:08.174595Z digest=sha256:c9bea15aa6dea9ee1e2731cb2cb0793c917a81c4f62292ae75e19ed6ab95540e

Observation a57cad7f-0a66-442d-ba67-07d698debdec · outbound

This paper cites Diverse and Effective Red Teaming with Auto-generated Rewards and Multi-step Reinforcement Learning.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Diverse and Effective Red Teaming with Auto-generated Rewards and Multi-step Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.215841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.215841Z digest=sha256:57f04d09a04bfabc06c959cece57cae431415061a9aac92a37faf0bc0516c0b3

Observation 0c83958e-51f1-4715-841a-38e611cdb273 · outbound

This paper cites Bhardwaj, D.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Bhardwaj, D

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:15.779030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:08.265880Z digest=sha256:b8cd41167cefe713a893f8f6d13a0ff4785d4a85ce8709ec7b8d593b4f2b7418

Observation bd7a3d64-e8c8-44d7-8978-302a0a69abae · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning On the Opportunities and Risks of Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.323501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.323501Z digest=sha256:c4a00ec8caffd912006c9dfcd48950bd003546da9c1ea6124fdc6f4015eb7946

Observation f393ed59-e8d9-45e9-8660-c820a41aa1e5 · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.370972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.370972Z digest=sha256:cf8086f7447d34e3e0bb1570aa858431b391f8894b6717564dbec32ac31f0d96

Observation b364c890-f672-4494-8c4e-0c3ca41f3b99 · outbound

This paper cites Chiang, L.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Chiang, L

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:15.678396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:08.431421Z digest=sha256:20d40153283529101e61cbab4db75e32940eb461f4e5bc6ba4333e6357f3555f

Observation 3a3d76af-8b39-4115-a880-59665351a6f2 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:15.529872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:08.467412Z digest=sha256:d8bdb4aa729027a27a853560eef8d2e764106ab2d1b390de5839e4a6df04ba3f

Observation 673ba450-cba6-4eeb-b443-628662efb98e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.506447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.506447Z digest=sha256:bd5ebd9f015354a40f7700291cf2b9a4369750ebdb4c8944c1bdc9211d4ff22e

Observation c18d2fbe-eeb3-4382-a8db-52aef4f2b1d5 · outbound

This paper cites The Llama 3 Herd of Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.542610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.542610Z digest=sha256:eda3ee810569b79f618078b126acb6de35ed713b2f4349df842f1af0f723f53c

Observation d985b050-1d04-41e4-95bc-c5fb7d70cc38 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:15.437821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:08.618088Z digest=sha256:c11203e8dc50f4df06f655531146793b5acfa14afc1ca6dbd108e570eec6b915

Observation 072fc78b-14de-47d8-9eb1-8f1452568d74 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:15.329250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:08.659593Z digest=sha256:851065ec35fbef0726f9c9892de4f13f98d3692f98526c32457caca36af3449e

Observation ca9c6f7a-6ddd-4ef9-9e2c-e833b8540891 · outbound

This paper cites Best-of-N Jailbreaking.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Best-of-N Jailbreaking

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.712974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.712974Z digest=sha256:43f18ff776a797a13c933de7793b14895aeb5ff677d98498d5dd4a878d2d010d

Observation eefbd577-524a-4cc1-a8d0-3faa2c8147af · outbound

This paper cites PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.749164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.749164Z digest=sha256:d59f0cb81fad590d49febc38864da5a8e64cb0d71166e39071065727266127cb

Observation f96588aa-b5a8-47c9-a3a3-77cb239334db · outbound

This paper cites Jiang, K.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Jiang, K

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:15.216923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:08.790854Z digest=sha256:2ca73d78e2a914bd4ae4a621ae93183763948487f3bf7a2b75d1f5ea68e67e26

Observation 8b7839bf-1144-4e59-9590-e49c9f2bbb38 · outbound

This paper cites Learning diverse attacks on large language models for robust red-teaming and safety tuning.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Learning diverse attacks on large language models for robust red-teaming and safety tuning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.855205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.855205Z digest=sha256:bc2931ff3be69fd095ce6d1b34780e817699ac7e53c0418b0449bb7afba652bd

Observation d9d7b985-ec8f-438f-b860-4a5163b078c8 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:15.081114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:08.904306Z digest=sha256:70043b06f692353abaa8bb818d54fa01906125fa02a0e23c84363dc4095ff1b7

Observation caa11580-6c71-4128-ad4f-253b6c47ab5d · outbound

This paper cites AmpleGCG: Learning a Universal and Transferable Generative Model of Adversarial Suffixes for Jailbreaking Both Open and Closed LLMs.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AmpleGCG: Learning a Universal and Transferable Generative Model of Adversarial Suffixes for Jailbreaking Both Open and Closed LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:08.949311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:08.949311Z digest=sha256:04328fe4cf9909eb03ba52e2a15c689a639b88bea590e9966948dc40db271262

Observation 33b04a5e-5201-4103-a505-f126f201da8a · outbound

This paper cites AutoDAN-Turbo: A Lifelong Agent for Strategy Self-Exploration to Jailbreak LLMs.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AutoDAN-Turbo: A Lifelong Agent for Strategy Self-Exploration to Jailbreak LLMs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.010611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.010611Z digest=sha256:84e1b4d275030cd5edabd0c1c823008abf17baa36a63c007b1b9801d12b3acfc

Observation 79503cc3-a8ae-4757-97cc-c009a116ff06 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:14.920262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:09.049209Z digest=sha256:c156fc1f10a5d68e4a50b32f6ac95866d20c0dcaa3c5822f118a8096e2cd90df

Observation 7ef22d1b-236d-4d76-b9a5-70094e50e563 · outbound

This paper cites Auto-RT: Automatic Jailbreak Strategy Exploration for Red-Teaming Large Language Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Auto-RT: Automatic Jailbreak Strategy Exploration for Red-Teaming Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.082636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.082636Z digest=sha256:43e781ae256da919eb9cf7fba08f191dc52ddb08f69cfd4be38e0b6dd552b3f7

Observation 9f76d43c-b567-41d5-8d90-ffe9b0a7e40a · outbound

This paper cites Mazeika, L.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Mazeika, L

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.799153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:09.159736Z digest=sha256:c9d6c68287e57a1aeab24e572bc1176a997fbd82e099482ebc27478bff513353

Observation b1e15a24-cb77-4191-9ab0-92c61ff84c58 · outbound

This paper cites Mehrotra, M.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Mehrotra, M

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.626933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:09.219879Z digest=sha256:aba4004ed6093b97126807ebe9086335733d9f9ad6c997fdc05d29d3ec472825

Observation 566053bd-80f3-4d37-853f-df3a49293b28 · outbound

This paper cites Gpt-3.5 turbo, 2023.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Gpt-3.5 turbo, 2023

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.498579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:09.278456Z digest=sha256:1eff2d26a42980547dc14cdc72390a693e4911c3d96c672089c1441c965841ba

Observation 626e1ff1-468b-4aaa-bfb7-e71a93d30c8f · outbound

This paper cites Gpt-4o system card, 2024a.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Gpt-4o system card, 2024a

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.309406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:09.326484Z digest=sha256:b12776e55da841e4791688d8de157fa1ff89a27887305005bd3ff0d6c158e312

Observation 08782bc4-4c05-447a-8f7a-e6054cf6024a · outbound

This paper cites AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.355215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.355215Z digest=sha256:63580e8ad345f8d217802241e968ce477f06b272bc42930b2d9cf27dbd78e938

Observation e8db39e9-155c-47b6-9fc9-841594d08b36 · outbound

This paper cites Perez, S.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Perez, S

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:14.151394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:09.433892Z digest=sha256:80dd81151e4aa89b62c6a356b7b5c625e4d17f52321710872ad9890d4b2cd766

Observation 8608203b-ab0d-4b7d-9e6d-5f5952b7c7aa · outbound

This paper cites ToolRL: Reward is All Tool Learning Needs.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning ToolRL: Reward is All Tool Learning Needs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.512168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.512168Z digest=sha256:7bf5ee6321cd516479da5eaf94f379ae54277ec573decf2fbf2409c18d88ddc1

Observation a61ec825-55bf-40eb-92e3-e9999b760f4e · outbound

This paper cites Samvelyan, S.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Samvelyan, S

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:13.886647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:09.600327Z digest=sha256:3346824e1db1ae0338171211182e18191c6ed27d458d6624e90b11a7c10061f9

Observation 2fe3d659-1d9e-4572-83d3-d16129c3df13 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.657922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.657922Z digest=sha256:ee1444c0d1027b20aee2f36986d3c920c6cfc74aea464092de0fdef6cc964c62

Observation 337364ca-5d04-4c9d-aa9e-140249eaeb30 · outbound

This paper cites Shaikh, H.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Shaikh, H

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:13.639232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:09.746566Z digest=sha256:3a9f947e99101507007ab2a03b7e75366436c1a404b449af9f44c74072c87212

Observation 33a0d048-074e-4b90-8607-1d344d8b45c1 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.833907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.833907Z digest=sha256:d7084fbb0dd05cb1fe78f775582571528e84c928ccb14b971cb3913770ffad4e

Observation 1da0f343-f8de-4521-ae3e-d911e0b7fecd · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:09.906691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:09.906691Z digest=sha256:476c46cc70e1d380c5478fd8f16f84a007aaba435afc20c4a4bb144ce1ca189e

Observation c80dcfb9-4c3d-4a0d-b3d0-46c2a6f5d20a · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.034542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.034542Z digest=sha256:8a6aed3009f08c4b436f81492dae0a83c517106c160a4df692263076c2e6a08e

Observation 20bb5f9f-9282-4eea-9926-6fe1ae134672 · outbound

This paper cites Wang and K.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Wang and K

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:13.402890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:10.174828Z digest=sha256:f72e828f4741a16391027a5a8611dd373810b20856e816be33f19ca292ace73b

Observation 9116b090-aef8-4210-95f5-bf89a5cdba32 · outbound

This paper cites Qwen2 Technical Report.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Qwen2 Technical Report

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.230800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.230800Z digest=sha256:ae179c35ba7196bf2165361436989ddad196f3449ab9a8831b8c47823148ceef

Observation 964183fa-aa52-49f2-927f-c5173047915a · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:13.185963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:10.304646Z digest=sha256:23dee8e56e9755c8d6b0ed41d964cf74cffb07836a107556e5119d0aa8bcd027

Observation d46e43ee-3901-486a-9347-9043cb9922c4 · outbound

This paper cites STAIR: Improving Safety Alignment with Introspective Reasoning.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning STAIR: Improving Safety Alignment with Introspective Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.437566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.437566Z digest=sha256:a1599769b3f408ddef783fe682413bc9ce85c6566226f14542553c849e022cc5

Observation 9be8f1ce-bb31-4409-9aea-d7454a72728f · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:13.013773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:10.548907Z digest=sha256:3f7bed4b826fec06037b231f652632918216bf296addd870c38cd26a699796db

Observation 0aeb746e-80aa-4229-83f4-d714ffbf2ddf · outbound

This paper cites Toward Optimal LLM Alignments Using Two-Player Games.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Toward Optimal LLM Alignments Using Two-Player Games

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.671950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.671950Z digest=sha256:facf1a39b9dfff9eb8a909b7bfc6c6fee99668b2160eb418c559f334bfa581c0

Observation 362e0e77-debc-4ddb-a8c9-f1c8227e1e3d · outbound

This paper cites AutoRedTeamer: Autonomous Red Teaming with Lifelong Attack Integration.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning AutoRedTeamer: Autonomous Red Teaming with Lifelong Attack Integration

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.771523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.771523Z digest=sha256:744a4bcc88a15e18a27e75a31c24ffc8138e741ea900e8c4d705c17c312add3f

Observation ed3ab69c-a3f5-45fb-a68a-4206b040c247 · outbound

This paper cites Purple-teaming LLMs with Adversarial Defender Training.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Purple-teaming LLMs with Adversarial Defender Training

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:10.911575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:10.911575Z digest=sha256:b74124a22bd80a4e6fe3274138aea4a093e66188b6562f1d9235c64f46d6a06f

Observation cc43c2c8-46e6-48aa-b081-3da843160ed5 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:11.128250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:11.128250Z digest=sha256:15e34c8d87865055f48d0c08c654d628873c2ff48204f9fece240b9d80817cfb

Observation 2d3ff433-976c-4255-ac9a-da42342f056e · outbound

This paper cites After obtaining the filtered 2k samples, we prompt the Qwen2.5-7B-Instruct model to imitate the sample attack as an example.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning After obtaining the filtered 2k samples, we prompt the Qwen2.5-7B-Instruct model to imitate the sample attack as an example

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:01:12.850546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:11.261804Z digest=sha256:8b0c73a96f6430b9a1aa489bdf0394f93c59804f19ae177f158cc6147f271d4d

Observation e1cf1b40-4758-4b2e-8d33-9e64c96a27d7 · outbound

This paper cites an unresolved cited work.

Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:01:12.612250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:01:11.473485Z digest=sha256:514518453a8728354d8476de2b750592ea2895bb25b202143e846b488de588c4

Pith citing papers

Observation 3caa8e0b-18b1-4a1e-a875-266a80aa3761 · inbound

Stable-GFlowNet: Toward Diverse and Robust LLM Red-Teaming via Contrastive Trajectory Balance cites this paper.

Stable-GFlowNet: Toward Diverse and Robust LLM Red-Teaming via Contrastive Trajectory Balance Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:41:16.865385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-09T19:29:36.348680Z digest=sha256:23a11f291ff03092735e12a3a2c851f832aec4385bd59c194522f6b3b110e3d6

Observation 4426f7f5-4c93-4191-8186-f553eb93f59c · inbound

Internalizing Safety Understanding in Large Reasoning Models via Verification cites this paper.

Internalizing Safety Understanding in Large Reasoning Models via Verification Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:51:14.269926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T01:50:59.283409Z digest=sha256:26f35252f4797c7093a237325ec58cbb325a5b85f3640423e6ce9afda5897e95

Observation 4645865c-d491-420f-917f-eb479a6b126e · inbound

Self-ReSET: Learning to Self-Recover from Unsafe Reasoning Trajectories cites this paper.

Self-ReSET: Learning to Self-Recover from Unsafe Reasoning Trajectories Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:11:16.194279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T02:07:04.364417Z digest=sha256:b721500833edcb17492791c55e5f44372a475839c1664afe1a92f075304fed49

Observation a31831b8-9265-4c4e-9e19-3609eaaf51d2 · inbound

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs cites this paper.

Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T04:16:34.711696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T09:21:57.373862Z digest=sha256:c305c97e69646b47d087f60b7ec6834035b5d459d80622f6e8ac224e548fdf5b

Observation e777a612-105b-4bf5-bd5a-b1cd2db8b8f5 · inbound

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents cites this paper.

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-11T22:40:37.839133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T22:40:37.839133Z digest=sha256:a471ee478db559f46cd1b936bb692354b236c664fac6eab42de55baa025c2088

Observation 0b78426c-7ced-42f2-b964-4f713e454373 · inbound

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents cites this paper.

Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-13T07:01:49.222325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T07:01:49.222325Z digest=sha256:bed60da66afa5dcd4f8e2d899655c64a57094fcb67294537d40973688c485727

Observation f8ab7de0-ef03-447a-9f60-7378374f6c54 · inbound

MJ: Multi-turn LLM Jailbreaking via Decomposed Credit Assignment cites this paper.

MJ: Multi-turn LLM Jailbreaking via Decomposed Credit Assignment Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-14T07:16:57.009797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:16:57.009797Z digest=sha256:7a9d61a5b64aab6d327844f069daed738809b49bafc5fac490aa37e44dec20e8

Observation 6b3f8d1d-b25c-4b2c-a058-d53d4c66aeff · inbound

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs cites this paper.

An Early Warning of Emerging Biosecurity Risks in Frontier LLMs Jailbreak-R1: Exploring the Jailbreak Capabilities of LLMs via Reinforcement Learning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T16:21:26.984437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:21:26.984437Z digest=sha256:38b1834d623dd0983250174aaa21ac0c289982301542f1914ae7d00e2fe2495e