Pith. sign in

Paper Citation Record · LEDGER

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs

As of 20 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 1 inbound Pith citation observation for arXiv:2504.19019.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.19019 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:08:38.276511Z

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T18:03:07.646917Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

26 of 26 outbound references displayed

  • verified exact1
  • verified fuzzy7
  • unresolved17
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 12949660-9994-4df5-b3dc-21c004e38d33 · outbound

This paper cites URL https://link.springer.com/article/10.1007/s11948-022-00364-7.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs URL https://link.springer.com/article/10.1007/s11948-022-00364-7

Reference 3

Resolution
verified exact
doi, observed 2026-08-16T10:08:38.461100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:08:37.952929Z digest=sha256:07055345b7c1713a376393e604d45bf965f1accd2c27bb6863a58573e9a13a82

Observation 467a8fbb-4af0-4768-831a-9229636e7171 · outbound

This paper cites doi: 10.18653/v1/2023.findings-emnlp.932.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2023.findings-emnlp.932

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:37.964360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:37.964360Z digest=sha256:ccc4ea31742710e0937db5f05e8628f1f1c01077a4e544ac9fd29d82cc5555a0

Observation 6cb19271-a66a-4edf-a2a3-1f6fcd6f4ce8 · outbound

This paper cites an unresolved cited work.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:08:38.920152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:08:37.969253Z digest=sha256:d50ba3ea061a438d773cbf3d7c6c5051491b0054a3a815e88e4980fd6552c9d4

Observation cd5250e7-33bd-4b42-850d-20dc21e73ca9 · outbound

This paper cites doi: 10.18653/v1/2023.acl-long.754.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2023.acl-long.754

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:37.976972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:37.976972Z digest=sha256:af917e8bd1353dfd1d596a419c7c6385a9df888d543b6bab57bf1b37f7e482ce

Observation 69d6aad1-3aae-4f74-9e71-3dc016ee9a8f · outbound

This paper cites Eric Wallace, Shi Feng, Nikhil Kandpal, Matt Gardner, and Sameer Singh.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Eric Wallace, Shi Feng, Nikhil Kandpal, Matt Gardner, and Sameer Singh

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.889780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:08:37.982898Z digest=sha256:68a74bd1deba5ce9125d1c8c95fefed40c3153869a787ce914a0645d8a26fd32

Observation 84e3cac8-4777-4ae9-92d8-4928c92545e0 · outbound

This paper cites doi: 10.18653/v1/D19-1461.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/D19-1461

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.098320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.098320Z digest=sha256:3708f393ec236c533d14ed14db88490fcb30c34189eee6039f44856904532596

Observation 39f9986c-e325-43a8-891a-2fbd9e8dfd3b · outbound

This paper cites doi: 10.18653/v1/2021.naacl-main.235.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2021.naacl-main.235

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.165350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.165350Z digest=sha256:3214fa8d3e856642938deaf117f5f4832cdbbcf4b5e38e538cbe2024bbab99d8

Observation 613e68f6-9ccb-472f-bdc1-c7ca97148e86 · outbound

This paper cites Reid Pryzant, Dan Iter, Jerry Li, Yin Lee, Chenguang Zhu, and Michael Zeng.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Reid Pryzant, Dan Iter, Jerry Li, Yin Lee, Chenguang Zhu, and Michael Zeng

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.853325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:08:38.169979Z digest=sha256:28e5ab2d3214b8f2c08c56efdb5558b63d84f451f0219d3ea3b177629ec6639a

Observation c531a573-ec15-4747-aa5a-0498e1b63ec7 · outbound

This paper cites doi: 10.18653/v1/2023.emnlp-main.494.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2023.emnlp-main.494

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.174154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.174154Z digest=sha256:aa686295f2aee75ed303111480527083d639c55afdc4b017c72301e58d7b0df6

Observation 3c201358-ebd2-436e-b9ed-1dbf97cfae98 · outbound

This paper cites doi: 10.18653/v1/2020.emnlp-main.346.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2020.emnlp-main.346

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.179479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.179479Z digest=sha256:0f9262af23627cfed19c735dfba015bd6b9edeb840311267eda22344bd841d35

Observation e2e520cc-375f-4f4c-ab1b-f429b589bdd8 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Chain-of-thought prompting elicits reasoning in large language models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.835801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:08:38.185102Z digest=sha256:8fddc3cc813fc97f8d1ec059cc3f4795b9076045421d29fb172c81f4fe873c7c

Observation bd3e3366-5e87-40f7-9ad9-8338084c147b · outbound

This paper cites Sean Wu, Michael Koo, Lesley Blum, Andy Black, Liyo Kao, Fabien Scalzo, and Ira Kurtz.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Sean Wu, Michael Koo, Lesley Blum, Andy Black, Liyo Kao, Fabien Scalzo, and Ira Kurtz

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.770663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:08:38.190481Z digest=sha256:1276aeaca1fa0a0fd62c5180d00ec3af966989bbb32a0f3f62791f9a07245c64

Observation c5ef1069-59aa-444d-b602-89dbe6be3b7d · outbound

This paper cites H-CoT: Hijacking the Chain-of-Thought Safety Reasoning Mechanism to Jailbreak Large Reasoning Models, Including OpenAI o1/o3, DeepSeek-R1, and Gemini 2.0 Flash Thinking.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs H-CoT: Hijacking the Chain-of-Thought Safety Reasoning Mechanism to Jailbreak Large Reasoning Models, Including OpenAI o1/o3, DeepSeek-R1, and Gemini 2.0 Flash Thinking

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.200391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.200391Z digest=sha256:1542a6df8866cd4a320f31cf79a126a201f728ce95337a526eb6338b13af2a69

Observation 2bedfcb5-7707-4855-baab-7b73aee2666a · outbound

This paper cites cc/paper_files/paper/2023/file/fd6613131889a4b656206c50a8bd7790-Paper-Conference.pdf.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs cc/paper_files/paper/2023/file/fd6613131889a4b656206c50a8bd7790-Paper-Conference.pdf

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.693167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:08:38.206366Z digest=sha256:aa21c556241b4f595371b03474ef0a8150f8ccd929997983d480b8ffca345bfd

Observation dc7bf599-49b2-4d19-8086-dda11735fdfc · outbound

This paper cites doi: 10.18653/v1/2024.acl-long.773.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2024.acl-long.773

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.211124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.211124Z digest=sha256:280b0fb5533c512261a62e141afc34744451d4d3bb39dcaf03042597b21580cd

Observation 07a4c418-9f55-4c5e-8237-cfdbbb686d77 · outbound

This paper cites Ethan Perez, Saffron Huang, Francis Song, Trevor Cai, Roman Ring, John Aslanides, Amelia Glaese, Nat McAleese, and Geoffrey Irving.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Ethan Perez, Saffron Huang, Francis Song, Trevor Cai, Roman Ring, John Aslanides, Amelia Glaese, Nat McAleese, and Geoffrey Irving

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.218933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.218933Z digest=sha256:9955f10284a42686e622850681d836e440f0a2d6dd8360885242b976b0d347aa

Observation e0c84e73-d367-408b-a63d-297f88c73d90 · outbound

This paper cites doi: 10.18653/v1/2022.emnlp-main.225.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2022.emnlp-main.225

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.224101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.224101Z digest=sha256:fd79a597c9cdcac646f419be411cbad3b6f84bcf13eff12a29558bede9200a7c

Observation c1a15d8d-f52f-4203-83ca-b45ca3452211 · outbound

This paper cites an unresolved cited work.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:08:38.675413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:08:38.276511Z digest=sha256:1a2186cdee0d1c1dc504362fe30c05f6d03d23fa8fbfd08821bdf5398ec06e29

Observation 8ec28ee2-c048-4408-8712-c8949403abfc · outbound

This paper cites doi: 10.18653/v1/P18-2006.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/P18-2006

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.228920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.228920Z digest=sha256:99ebe32905a2d30a4211890ece0a34894b0f3d9c0519f4e473cbd5a5387633cf

Observation 47e429f2-35ff-4f38-a9cb-b8aa2cd2f583 · outbound

This paper cites doi: 10.18653/v1/D19-1221.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/D19-1221

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:37.988832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:37.988832Z digest=sha256:74ea73e1c71c4fbfc24467d8e7676a1ac02d0b48ca024273a5d8a9c8b9ccda96

Observation 39f32b0d-d0b2-42e1-8a86-c151b11568f3 · outbound

This paper cites doi: 10.18653/v1/2020.acl-main.442.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/2020.acl-main.442

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.141110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.141110Z digest=sha256:676da4fb7ff44f0f9867c09040a4d72261e06a95558da32fbdb4f2ea286f7c63

Observation 411f126d-d1d2-4013-b57d-f7f466f1e3c4 · outbound

This paper cites press/v139/leino21a.html.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs press/v139/leino21a.html

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.873502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:08:38.036358Z digest=sha256:85a8adebfaf676189c45f360e05da44d140cd2916e12333b61332242f4beb549

Observation 03f11b64-972f-44fd-acd3-fd54b741ed2c · outbound

This paper cites doi: 10.18653/v1/ 2022.acl-short.94.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs doi: 10.18653/v1/ 2022.acl-short.94

Reference 2022

Resolution
malformed identifier
no resolver link, observed 2026-08-16T10:08:37.947146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:37.947146Z digest=sha256:67fe7be2b4798647930044bb1a57b70d79668a93adf7480cb22491658e602ca6

Observation 7a21dce0-1890-4913-ac7d-7320ccc2a3b8 · outbound

This paper cites Thorny roses: Investigating the dual use dilemma in natural language processing.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Thorny roses: Investigating the dual use dilemma in natural language processing

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:08:38.953818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:08:37.958667Z digest=sha256:cb21921d6e614591e31f3d5d12b9e0925b643c9b344ef13c5d0da840620d487a

Observation 54316ddb-ab7c-4826-bcfb-76d02dcb0334 · outbound

This paper cites URL https://ojs.aaai.org/index.php/AAAI/ article/view/29720.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs URL https://ojs.aaai.org/index.php/AAAI/ article/view/29720

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:37.855493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:37.855493Z digest=sha256:47f6de9918ded009833e743aba19e6076f46857bd8458733a2cecdb3d1b11877

Observation a0772dce-cafd-4a26-96b6-c694a7ee1034 · outbound

This paper cites Martin Kuo, Jianyi Zhang, Aolin Ding, Qinsi Wang, Louis DiValentin, Yujia Bao, Wei Wei, Hai Li, and Yiran Chen.

Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs Martin Kuo, Jianyi Zhang, Aolin Ding, Qinsi Wang, Louis DiValentin, Yujia Bao, Wei Wei, Hai Li, and Yiran Chen

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-16T10:08:38.195470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:08:38.195470Z digest=sha256:b1e19a62c09fb1ab9039f6a8af82d809538ad2c4bfda710d73ed3e7b45a4e1f8

Pith citing papers

Observation c8671530-3e35-41c0-90a5-b96d53df5514 · inbound

PQR: A Framework to Generate Diverse and Realistic User Queries that Elicit QA Agent Failures cites this paper.

PQR: A Framework to Generate Diverse and Realistic User Queries that Elicit QA Agent Failures Graph of Attacks: Improved Black-Box and Interpretable Jailbreaks for LLMs

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-20T18:03:36.776688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-20T18:03:07.646917Z digest=sha256:0e3e4a3882e256f23e0ae612382469497e961c397d1bea6a8d4c7e4b88c91ab5