Pith. sign in

Paper Citation Record · LEDGER

RL-JACK: Reinforcement Learning-powered Black-box Jailbreaking Attack against LLMs

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2406.08725.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.08725 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:07:02.583100Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T13:35:46.523325Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 98a71be8-0944-4049-92ef-4250b9fb730f · inbound

VERA: Variational Inference Framework for Jailbreaking Large Language Models cites this paper.

VERA: Variational Inference Framework for Jailbreaking Large Language Models RL-JACK: Reinforcement Learning-powered Black-box Jailbreaking Attack against LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T22:07:02.583100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:07:02.583100Z digest=sha256:b04f2f935899c9e0b2f6b4fcca3be6003703653c5dc21bbfa5b4c607c83e6821

Observation 0a5f4e48-9e37-4a63-a0eb-924e89f6fbe0 · inbound

MGC: A Compiler Framework Exploiting Compositional Blindness in Aligned LLMs for Malware Generation cites this paper.

MGC: A Compiler Framework Exploiting Compositional Blindness in Aligned LLMs for Malware Generation RL-JACK: Reinforcement Learning-powered Black-box Jailbreaking Attack against LLMs

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T20:45:53.261193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:45:53.261193Z digest=sha256:fe649d8e63ed4c4116ce6425777d5655ba1c540b0107250330579c3d71a4a37c

Observation 097a2d0e-b2dc-4b40-8782-bf15bc108d3d · inbound

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM cites this paper.

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM RL-JACK: Reinforcement Learning-powered Black-box Jailbreaking Attack against LLMs

Reference 132

Resolution
unresolved
no resolver link, observed 2026-08-05T23:13:04.875158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:13:04.875158Z digest=sha256:daf35194724190f369a57c64d00beb6552a3948944a4cd828759db756b7b614a

Observation 83850c27-a627-4be6-86ff-bd3abf98e28d · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses RL-JACK: Reinforcement Learning-powered Black-box Jailbreaking Attack against LLMs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:38.012962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:38.012962Z digest=sha256:8eba804ac44d307e750c40b901857eec7ef6be1b54d0c02f3fcb6376b890a8a5

Observation d3eadf7c-d305-48c5-8a7e-1c04d409b8ab · inbound

A Systematic Investigation of RL-Jailbreaking in LLMs cites this paper.

A Systematic Investigation of RL-Jailbreaking in LLMs RL-JACK: Reinforcement Learning-powered Black-box Jailbreaking Attack against LLMs

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:40:52.616612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:32:42.151644Z digest=sha256:8007ed5dd5d3ae6fb785d9fb2134c525745892d4d859c8ac0da346fceca3046b

Observation 680dcfc8-2598-44e1-8abe-cb4c06169eed · inbound

A Systematic Investigation of RL-Jailbreaking in LLMs cites this paper.

A Systematic Investigation of RL-Jailbreaking in LLMs RL-JACK: Reinforcement Learning-powered Black-box Jailbreaking Attack against LLMs

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:35:46.524911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T22:59:07.861941Z digest=sha256:36d8d220f5753dfca93279695cfa2ce2e45bed917fc46c0b38839aee091b62ba

Observation 5bd37c73-d0de-4eec-9211-ad67178d0348 · inbound

A Systematic Investigation of RL-Jailbreaking in LLMs cites this paper.

A Systematic Investigation of RL-Jailbreaking in LLMs RL-JACK: Reinforcement Learning-powered Black-box Jailbreaking Attack against LLMs

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-02T14:41:05.948928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:41:05.948928Z digest=sha256:3d0662f6df5344439d75e2cd847c3f25ac56cbf840016b1d0a8214701489311a