Pith. sign in

Paper Citation Record · LEDGER

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs

As of 20 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 1 inbound Pith citation observation for arXiv:2504.19019.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.19019 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:08:38.276511Z

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T18:03:07.646917Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

26 of 26 outbound references displayed

  • verified exact1
  • verified fuzzy7
  • unresolved17
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 12949660-9994-4df5-b3dc-21c004e38d33 · outbound

This paper cites URL https://link.springer.com/article/10.1007/s11948-022-00364-7.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs URL https://link.springer.com/article/10.1007/s11948-022-00364-7

Reference 3

Resolution
verified exact
doi, observed 2026-08-16T10:08:38.461100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:08:37.952929Z digest=sha256:1f7ccbf9064f419b82ce794139e3cfa09ed1cb625ddbb2ca5e3870bea9be42f0

Observation 467a8fbb-4af0-4768-831a-9229636e7171 · outbound

This paper cites doi: 10.18653/v1/2023.findings-emnlp.932.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2023.findings-emnlp.932

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:37.964360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:37.964360Z digest=sha256:ccc4ea31742710e0937db5f05e8628f1f1c01077a4e544ac9fd29d82cc5555a0

Observation 6cb19271-a66a-4edf-a2a3-1f6fcd6f4ce8 · outbound

This paper cites an unresolved cited work.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:08:38.920152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:08:37.969253Z digest=sha256:f49bfb664f13c8f90fed240f25abc6ee90773eafce5b3e316dea17e98968b6fe

Observation cd5250e7-33bd-4b42-850d-20dc21e73ca9 · outbound

This paper cites doi: 10.18653/v1/2023.acl-long.754.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2023.acl-long.754

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:37.976972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:37.976972Z digest=sha256:af917e8bd1353dfd1d596a419c7c6385a9df888d543b6bab57bf1b37f7e482ce

Observation 69d6aad1-3aae-4f74-9e71-3dc016ee9a8f · outbound

This paper cites Eric Wallace, Shi Feng, Nikhil Kandpal, Matt Gardner, and Sameer Singh.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Eric Wallace, Shi Feng, Nikhil Kandpal, Matt Gardner, and Sameer Singh

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.889780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:08:37.982898Z digest=sha256:89d24f2fe1581e7e04db6b720298d26755cf464b2befb3724a6b52473af68894

Observation 84e3cac8-4777-4ae9-92d8-4928c92545e0 · outbound

This paper cites doi: 10.18653/v1/D19-1461.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/D19-1461

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.098320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.098320Z digest=sha256:3708f393ec236c533d14ed14db88490fcb30c34189eee6039f44856904532596

Observation 39f9986c-e325-43a8-891a-2fbd9e8dfd3b · outbound

This paper cites doi: 10.18653/v1/2021.naacl-main.235.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2021.naacl-main.235

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.165350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.165350Z digest=sha256:3214fa8d3e856642938deaf117f5f4832cdbbcf4b5e38e538cbe2024bbab99d8

Observation 613e68f6-9ccb-472f-bdc1-c7ca97148e86 · outbound

This paper cites Reid Pryzant, Dan Iter, Jerry Li, Yin Lee, Chenguang Zhu, and Michael Zeng.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Reid Pryzant, Dan Iter, Jerry Li, Yin Lee, Chenguang Zhu, and Michael Zeng

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.853325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:08:38.169979Z digest=sha256:ac2c486ef52b40168869bba6ba47c6ea295b6ff377af662aff9375f1c054908f

Observation c531a573-ec15-4747-aa5a-0498e1b63ec7 · outbound

This paper cites doi: 10.18653/v1/2023.emnlp-main.494.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2023.emnlp-main.494

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.174154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.174154Z digest=sha256:aa686295f2aee75ed303111480527083d639c55afdc4b017c72301e58d7b0df6

Observation 3c201358-ebd2-436e-b9ed-1dbf97cfae98 · outbound

This paper cites doi: 10.18653/v1/2020.emnlp-main.346.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2020.emnlp-main.346

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.179479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.179479Z digest=sha256:0f9262af23627cfed19c735dfba015bd6b9edeb840311267eda22344bd841d35

Observation e2e520cc-375f-4f4c-ab1b-f429b589bdd8 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Chain-of-thought prompting elicits reasoning in large language models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.835801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:08:38.185102Z digest=sha256:1c71ad60db4a40b658d69e8f4307437ba5a8d4a92be2902869fea6141a35161a

Observation bd3e3366-5e87-40f7-9ad9-8338084c147b · outbound

This paper cites Sean Wu, Michael Koo, Lesley Blum, Andy Black, Liyo Kao, Fabien Scalzo, and Ira Kurtz.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Sean Wu, Michael Koo, Lesley Blum, Andy Black, Liyo Kao, Fabien Scalzo, and Ira Kurtz

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.770663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:08:38.190481Z digest=sha256:a79400084c914e342fd4b2d5482b26fdfffc18da7ec597da0a5eac1b5f20e8a9

Observation c5ef1069-59aa-444d-b602-89dbe6be3b7d · outbound

This paper cites H-CoT: Hijacking the Chain-of-Thought Safety Reasoning Mechanism to Jailbreak Large Reasoning Models, Including OpenAI o1/o3, DeepSeek-R1, and Gemini 2.0 Flash Thinking.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs H-CoT: Hijacking the Chain-of-Thought Safety Reasoning Mechanism to Jailbreak Large Reasoning Models, Including OpenAI o1/o3, DeepSeek-R1, and Gemini 2.0 Flash Thinking

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.200391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.200391Z digest=sha256:1542a6df8866cd4a320f31cf79a126a201f728ce95337a526eb6338b13af2a69

Observation 2bedfcb5-7707-4855-baab-7b73aee2666a · outbound

This paper cites cc/paper_files/paper/2023/file/fd6613131889a4b656206c50a8bd7790-Paper-Conference.pdf.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs cc/paper_files/paper/2023/file/fd6613131889a4b656206c50a8bd7790-Paper-Conference.pdf

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.693167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:08:38.206366Z digest=sha256:32f1e1ec6875ef17e217fd39e0d704772c04cae7de2fbb21277c58c73bb296b7

Observation dc7bf599-49b2-4d19-8086-dda11735fdfc · outbound

This paper cites doi: 10.18653/v1/2024.acl-long.773.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2024.acl-long.773

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.211124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.211124Z digest=sha256:280b0fb5533c512261a62e141afc34744451d4d3bb39dcaf03042597b21580cd

Observation 07a4c418-9f55-4c5e-8237-cfdbbb686d77 · outbound

This paper cites Ethan Perez, Saffron Huang, Francis Song, Trevor Cai, Roman Ring, John Aslanides, Amelia Glaese, Nat McAleese, and Geoffrey Irving.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Ethan Perez, Saffron Huang, Francis Song, Trevor Cai, Roman Ring, John Aslanides, Amelia Glaese, Nat McAleese, and Geoffrey Irving

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.218933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.218933Z digest=sha256:9955f10284a42686e622850681d836e440f0a2d6dd8360885242b976b0d347aa

Observation e0c84e73-d367-408b-a63d-297f88c73d90 · outbound

This paper cites doi: 10.18653/v1/2022.emnlp-main.225.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2022.emnlp-main.225

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.224101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.224101Z digest=sha256:fd79a597c9cdcac646f419be411cbad3b6f84bcf13eff12a29558bede9200a7c

Observation c1a15d8d-f52f-4203-83ca-b45ca3452211 · outbound

This paper cites an unresolved cited work.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:08:38.675413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:08:38.276511Z digest=sha256:16950e22b6d14876b337087dc47c2ed87d9844a25f18e35dbe199e98b0ef2de0

Observation 8ec28ee2-c048-4408-8712-c8949403abfc · outbound

This paper cites doi: 10.18653/v1/P18-2006.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/P18-2006

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.228920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.228920Z digest=sha256:99ebe32905a2d30a4211890ece0a34894b0f3d9c0519f4e473cbd5a5387633cf

Observation 47e429f2-35ff-4f38-a9cb-b8aa2cd2f583 · outbound

This paper cites doi: 10.18653/v1/D19-1221.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/D19-1221

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:37.988832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:37.988832Z digest=sha256:74ea73e1c71c4fbfc24467d8e7676a1ac02d0b48ca024273a5d8a9c8b9ccda96

Observation 39f32b0d-d0b2-42e1-8a86-c151b11568f3 · outbound

This paper cites doi: 10.18653/v1/2020.acl-main.442.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2020.acl-main.442

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.141110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.141110Z digest=sha256:676da4fb7ff44f0f9867c09040a4d72261e06a95558da32fbdb4f2ea286f7c63

Observation 411f126d-d1d2-4013-b57d-f7f466f1e3c4 · outbound

This paper cites press/v139/leino21a.html.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs press/v139/leino21a.html

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.873502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:08:38.036358Z digest=sha256:6c4b7264fc9afdd7c72cc9e022d9eca563b241eeb59f7dbab9cc294c7aa9b556

Observation 03f11b64-972f-44fd-acd3-fd54b741ed2c · outbound

This paper cites doi: 10.18653/v1/ 2022.acl-short.94.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/ 2022.acl-short.94

Reference 2022

Resolution
malformed identifier
no resolver link, observed 2026-08-16T10:08:37.947146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:37.947146Z digest=sha256:67fe7be2b4798647930044bb1a57b70d79668a93adf7480cb22491658e602ca6

Observation 7a21dce0-1890-4913-ac7d-7320ccc2a3b8 · outbound

This paper cites Thorny roses: Investigating the dual use dilemma in natural language processing.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Thorny roses: Investigating the dual use dilemma in natural language processing

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.953818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T10:08:37.958667Z digest=sha256:f22d3b9439e4f8b65c347c1487f6ee5ce56898fec48c2e53392c11107a1eda10

Observation 54316ddb-ab7c-4826-bcfb-76d02dcb0334 · outbound

This paper cites URL https://ojs.aaai.org/index.php/AAAI/ article/view/29720.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs URL https://ojs.aaai.org/index.php/AAAI/ article/view/29720

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:37.855493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:37.855493Z digest=sha256:47f6de9918ded009833e743aba19e6076f46857bd8458733a2cecdb3d1b11877

Observation a0772dce-cafd-4a26-96b6-c694a7ee1034 · outbound

This paper cites Martin Kuo, Jianyi Zhang, Aolin Ding, Qinsi Wang, Louis DiValentin, Yujia Bao, Wei Wei, Hai Li, and Yiran Chen.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Martin Kuo, Jianyi Zhang, Aolin Ding, Qinsi Wang, Louis DiValentin, Yujia Bao, Wei Wei, Hai Li, and Yiran Chen

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.195470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.195470Z digest=sha256:b1e19a62c09fb1ab9039f6a8af82d809538ad2c4bfda710d73ed3e7b45a4e1f8

Pith citing papers

Observation c8671530-3e35-41c0-90a5-b96d53df5514 · inbound

PQR: A Framework to Generate Diverse and Realistic User Queries that Elicit QA Agent Failures cites this paper.

PQR: A Framework to Generate Diverse and Realistic User Queries that Elicit QA Agent Failures Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:03:36.776688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-20T18:03:07.646917Z digest=sha256:5f4991c04068e9818ced7785f565b94e80f4f652f8a973bde6b419980d2b90b0