Pith. sign in

Paper Citation Record · LEDGER

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models

As of 7 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 0 inbound Pith citation observations for arXiv:2506.18543.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.18543 v2

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:20:59.199192Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

68 of 68 outbound references displayed

  • verified exact1
  • verified fuzzy22
  • unresolved42
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation aed38e89-35de-4a6d-921c-1a46b2b99efa · outbound

This paper cites Improving language understanding by generative pre-training2018.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Improving language understanding by generative pre-training2018

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:04.904571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:51.740664Z digest=sha256:20e44c4c7d0fb43e2b86155ce99ab55b016135d2680010a0232df4498d01094e

Observation 1ef3f846-4409-4170-a3c4-4c062658fd0d · outbound

This paper cites GPT-4 Technical Report.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:51.894501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:51.894501Z digest=sha256:cfce60f2b3780a292d3dc72118dda07c79be7ceee4daea62ec74e00ecac14728

Observation d3dd119b-d654-4d6b-b168-1bdb893da6c2 · outbound

This paper cites DeepSeek LLM: Scaling Open-Source Language Models with Longtermism.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:51.997871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:51.997871Z digest=sha256:cf5897a707cde318d3b273657dca3172c740272842540b5b5e6a90d93f1f0320

Observation 475ed272-df6e-433f-8369-58ebefe17c1a · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:52.145353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:52.145353Z digest=sha256:09a052904965faf955b83402ad571765dbe155ed8f281331e17ddba0ac53bd17

Observation 2f51017b-8333-451b-9117-773d2e2a4d40 · outbound

This paper cites PAL: Proxy-Guided Black-Box Attack on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:52.222527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:52.222527Z digest=sha256:c79dec345d81347b3137058621190ba299441af0eca47a835303be4761bb0ef9

Observation 397bd4f2-0c8f-40b2-bb8e-1e50f9542ab5 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:52.358358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:52.358358Z digest=sha256:28acee7683a61969c708813288530ed94ef45e07b2a7e415199d5dd9558424b9

Observation cc1567fc-3ba2-4fc9-9c42-1ccde88ab3d8 · outbound

This paper cites Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:52.542586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:52.542586Z digest=sha256:1d0363bfcdd3636c08b494f9dac7a402c50453eb4bb4393d65b9deb9e483e0b4

Observation 3cf6324c-550c-4d8a-8cbb-105e3cb79ade · outbound

This paper cites A survey of backdoor attacks and defenses on large language models: Implications for security measures.Authorea Preprints2024.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models A survey of backdoor attacks and defenses on large language models: Implications for security measures.Authorea Preprints2024

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:04.636289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:52.790543Z digest=sha256:9ebcf83d97f07a6c5e887252971066d7cf0659d5c3b8d74024c429a88d273aa6

Observation b8189ea6-8014-43c2-9a13-f58b06787a87 · outbound

This paper cites Training language models to follow instructions with human feedback.NeurIPS2022,35, 27730–27744.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Training language models to follow instructions with human feedback.NeurIPS2022,35, 27730–27744

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:04.274233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:52.876347Z digest=sha256:95536a853d1282c6817fadc5323dc967ccc471242cabc862fe6fcda8b95e0212

Observation c6aa9a52-c618-443e-b245-10472b891979 · outbound

This paper cites Defending Large Language Models Against Jailbreak Attacks Through Chain of Thought Prompting.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Defending Large Language Models Against Jailbreak Attacks Through Chain of Thought Prompting

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:04.119138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:53.014644Z digest=sha256:374da15e527b7f2acdabd2c034398873d7ecf2441700b0c904c8a9be376469dd

Observation ad0c9440-6356-4961-8102-c23bd7bf22f1 · outbound

This paper cites Many-shot jailbreaking.NeurIPS2024,37, 129696–129742.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Many-shot jailbreaking.NeurIPS2024,37, 129696–129742

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:03.956256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:53.104384Z digest=sha256:690c048da068c571711987146b6fc6e6b7fc7f6d7a4afc4f73fe7c8c4b2f54ce

Observation 1107255d-f51b-401a-841e-7976b9a6d564 · outbound

This paper cites Pandora: Jailbreak GPTs by Retrieval Augmented Generation Poisoning.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Pandora: Jailbreak GPTs by Retrieval Augmented Generation Poisoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:53.234749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:53.234749Z digest=sha256:915b22b44f647b20695a3ebfd30d45daca2d54086ce4868429389907188aac74

Observation 2ec64b4f-08f6-42c6-a595-c514a2f21e81 · outbound

This paper cites Multi-step Jailbreaking Privacy Attacks on ChatGPT.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Multi-step Jailbreaking Privacy Attacks on ChatGPT

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:53.545039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:53.545039Z digest=sha256:eafab8df3703ef20087863937bcb6a70e40ce7edd0aecb05c4b7d27bcf122d8b

Observation 5b09016f-1598-49f5-8296-9045b92df4b2 · outbound

This paper cites Adversarial Demonstration Attacks on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Adversarial Demonstration Attacks on Large Language Models

Reference 14

Resolution
malformed identifier
no resolver link, observed 2026-08-06T23:20:53.674318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:53.674318Z digest=sha256:d71a71d43c5c2980d50cb27fe0f5be40a7bfb300ee5ab4ca1c2a238928ac66af

Observation acd11a48-ed3f-4c4d-be94-32473105ddd4 · outbound

This paper cites Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:53.797609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:53.797609Z digest=sha256:d9b258bf789a2b998cbc55d2f974d3b5637f544a24add15d058bf709fe238fc9

Observation 8c70371a-3c8c-4c98-8d98-02c6c726e70b · outbound

This paper cites Improved few-shot jailbreaking can circumvent aligned language models and their defenses.NeurIPS2024,37, 32856–32887.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Improved few-shot jailbreaking can circumvent aligned language models and their defenses.NeurIPS2024,37, 32856–32887

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:03.744360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:53.856016Z digest=sha256:ed42212e557bb74806065a2a9c9c53ba235fb976e1a660f6162e1d3c833da2e4

Observation c19d5bb5-90b4-41b0-a6fe-a093ffc61213 · outbound

This paper cites Play Guessing Game with LLM: Indirect Jailbreak Attack with Implicit Clues.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Play Guessing Game with LLM: Indirect Jailbreak Attack with Implicit Clues

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:53.974751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:53.974751Z digest=sha256:d1b59ec04f54be792ed5844d9627f0fcf4dd8a26df6cd2bb222ab3a3a54f06f5

Observation abd74381-d8d7-4181-8228-b93e0156df97 · outbound

This paper cites Artprompt: Ascii art-based jailbreak attacks against aligned llms.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Artprompt: Ascii art-based jailbreak attacks against aligned llms

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:03.585017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:54.051307Z digest=sha256:7e679197fd6588fcfdd15343d0ab8bfb0220befd76beb1073385d303a698fa2f

Observation 7934abf5-bb74-4271-a43f-dfa0858bc75c · outbound

This paper cites Semantic Mirror Jailbreak: Genetic Algorithm Based Jailbreak Prompts Against Open-source LLMs.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Semantic Mirror Jailbreak: Genetic Algorithm Based Jailbreak Prompts Against Open-source LLMs

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:54.194806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:54.194806Z digest=sha256:8c347f4738fc32422f838f9b6f9f152ccae1dbbeab653340ccef7e3c02111f01

Observation 62c6b020-c4b5-4cfb-ba2b-4431d76ecd1d · outbound

This paper cites Understanding and Enhancing the Transferability of Jailbreaking Attacks.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Understanding and Enhancing the Transferability of Jailbreaking Attacks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:54.264785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:54.264785Z digest=sha256:61b9ac0820a83c16e8cedf4180a7e68089593688d85346a0083fecdb55afcfde

Observation 04bf703d-e726-4859-8139-7bd43ff41a54 · outbound

This paper cites AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:54.385546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:54.385546Z digest=sha256:63ff9b9df7849a3c293e2fca1c542d66d07f6d749540b2a9ced7b0714cc683fd

Observation f737caa1-dd19-4063-b49f-336cf5337eaf · outbound

This paper cites Flipattack: Jailbreak llms via flipping.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Flipattack: Jailbreak llms via flipping

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:03.365496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:54.467206Z digest=sha256:fe991e0544790f5dba26a9720e572d5c7fd52a5c92dbafbc3b351629c6be9bfa

Observation b80279ff-90d9-451b-bac6-6d4c9d7bfbff · outbound

This paper cites All in how you ask for it: Simple black-box method for jailbreak attacks.Applied Sciences2024,14, 3558.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models All in how you ask for it: Simple black-box method for jailbreak attacks.Applied Sciences2024,14, 3558

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:03.130334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:54.614751Z digest=sha256:941348a8cb45c39ccb487cea860b9f8058908c38eb41508ef0d2139eb797a08e

Observation 66031097-22d5-4869-95db-3bd8274d5547 · outbound

This paper cites Emoji Attack: Enhancing Jailbreak Attacks Against Judge LLM Detection.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Emoji Attack: Enhancing Jailbreak Attacks Against Judge LLM Detection

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:02.924635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:54.718877Z digest=sha256:65597f261d664dc65d5a28893bb47ea9a9bfc2c3995cb9254829c3b7a768b1a8

Observation 572595dc-71e9-4f77-ab83-eafeb7994cbd · outbound

This paper cites The Dark Side of Trust: Authority Citation-Driven Jailbreak Attacks on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models The Dark Side of Trust: Authority Citation-Driven Jailbreak Attacks on Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:54.845285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:54.845285Z digest=sha256:60de2b746d930428ef075076ceedcc0d85ae7920bdc719963bcece6c0b2ff0bf

Observation 80a5b13b-f9b0-4d0d-8bd2-195e885fbab4 · outbound

This paper cites GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:54.935278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:54.935278Z digest=sha256:7f71abd0f28abc9c5ca1287accae46bd6f50acf0b53db088103a42bdc6a21f43

Observation fe8b4cf5-c9ed-4ddc-9b02-460cffe25f72 · outbound

This paper cites Gpt-4 is too smart to be safe: Stealthy chat with llms via cipher2024.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Gpt-4 is too smart to be safe: Stealthy chat with llms via cipher2024

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:02.700975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:55.064753Z digest=sha256:3244010dccf4e56b1bccac318b18173378124c823c167e1358cd4bbd80b4df63

Observation 54a66195-b915-43e4-9959-84fd95786d5c · outbound

This paper cites Explore, Establish, Exploit: Red Teaming Language Models from Scratch.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Explore, Establish, Exploit: Red Teaming Language Models from Scratch

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:55.144189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:55.144189Z digest=sha256:9512aaa4ce0ada8a577d269ffcd3cc4388a1ac8977ccccc3ab6e1ad288b2139a

Observation c04a6230-7c1e-48d7-a277-2891647b17b5 · outbound

This paper cites Jailbreaking black box large language models in twenty queries.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Jailbreaking black box large language models in twenty queries

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:02.554468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:55.255225Z digest=sha256:fef04a56b9fe4324f774fe8d7ac91657cf6f7f1e1754d0db5916d2c7b4e661ce

Observation 4e48964a-e6cf-40b9-95be-f7f0537b1c4b · outbound

This paper cites MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:55.350714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:55.350714Z digest=sha256:2077997ac9aca77dcb036164558871f2c029e808e481da29f6655702c65bb675

Observation 55094c79-bba4-4d98-b270-58aefc9fa2a9 · outbound

This paper cites MART: Improving LLM Safety with Multi-round Automatic Red-Teaming.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:02.386229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:55.425537Z digest=sha256:10a3df456135575488cf31940e63ab81f835da77ed0fde295766e6e32b95a221

Observation 4c262b5e-e8cb-4176-8743-7c89d623960d · outbound

This paper cites Guard: Role-playing to generate natural-language jailbreakings to test guideline adherence of large language models.arXiv preprint arXiv:2402.032992024.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Guard: Role-playing to generate natural-language jailbreakings to test guideline adherence of large language models.arXiv preprint arXiv:2402.032992024

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:55.484326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:55.484326Z digest=sha256:51df4c261eb18e4083948e6e31ef19667391196f0e34f8a51d8e190a37860b31

Observation 1fc88740-52ae-418d-8b91-0abd30c85890 · outbound

This paper cites Goal-Oriented Prompt Attack and Safety Evaluation for LLMs.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Goal-Oriented Prompt Attack and Safety Evaluation for LLMs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:55.602649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:55.602649Z digest=sha256:0132b306becb33bc80e0465c48b556e14611bc886eac18c724b6ef2d195200a8

Observation 4ed0e8b6-e09e-46ab-8341-c7523c28610e · outbound

This paper cites Tree of attacks: Jailbreaking black-box llms automatically.NeurIPS2024,37, 61065–61105.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Tree of attacks: Jailbreaking black-box llms automatically.NeurIPS2024,37, 61065–61105

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:02.212017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:55.752634Z digest=sha256:904d0f934430a97dbd8fdd71f634c0c987d54822a223bbf6eadc72672bb5869e

Observation 6d906230-1a4e-4b3f-b9a8-854bff233fbf · outbound

This paper cites Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:55.859123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:55.859123Z digest=sha256:f3faf8219b9021723244dad8d9d252e87cc5c314e7ca38a8761d265d3598e9b2

Observation 207ab80e-bb35-4fe7-b1ba-5e6a8bd3c6af · outbound

This paper cites Evil Geniuses: Delving into the Safety of LLM-based Agents.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Evil Geniuses: Delving into the Safety of LLM-based Agents

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:55.953808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:55.953808Z digest=sha256:73b93e771749cfc29f61c12936c44d557fde365a981cfef315368b20e1d99702

Observation 898e97be-5c05-4f88-9ce5-c472c57af4b6 · outbound

This paper cites How johnny can persuade llms to jailbreak them: Rethinking persuasion to challenge ai safety by humanizing llms.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models How johnny can persuade llms to jailbreak them: Rethinking persuasion to challenge ai safety by humanizing llms

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:02.015166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:56.055295Z digest=sha256:dfed07e5446931adc2822687f9bc496ef1dad76aaa25fbc17af66b38bd7ce21e

Observation 09cebd2c-7789-45e8-89f2-5eb8bf3a0b10 · outbound

This paper cites Attacking Large Language Models with Projected Gradient Descent.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Attacking Large Language Models with Projected Gradient Descent

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:56.204745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:56.204745Z digest=sha256:99a787ef3510444ec883dd42b09bbe68968b03fe111f32fe7c99f35807a2570c

Observation a469243d-98ab-448f-a553-d9a672b4148f · outbound

This paper cites Query-based adversarial prompt generation.NeurIPS2024, 37, 128260–128279.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Query-based adversarial prompt generation.NeurIPS2024, 37, 128260–128279

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:01.842246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:56.319414Z digest=sha256:70225455faba004856ea114769b4702f3aee587d0d4327c6d2629cc5ed3a8df8

Observation d2bfecdd-1356-4824-8494-fac5065e7174 · outbound

This paper cites Improved Techniques for Optimization-Based Jailbreaking on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Improved Techniques for Optimization-Based Jailbreaking on Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:56.482120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:56.482120Z digest=sha256:685e42c2efa810e830c99aa264f6dcbed9f3644ddfccbbbbeb639e1377261039

Observation 2309b2d9-7282-4885-9774-dc260c466e86 · outbound

This paper cites Iterative self-tuning llms for enhanced jailbreaking capabilities.arXiv preprint arXiv:2410.184692024.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Iterative self-tuning llms for enhanced jailbreaking capabilities.arXiv preprint arXiv:2410.184692024

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:56.584927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:56.584927Z digest=sha256:9cb15c8cc35cc3aa890a38a79198a0337623fa12b7d28ba56dda6d463826db62

Observation 980df864-1235-45f1-a729-c8df39e53787 · outbound

This paper cites From noise to clarity: Unraveling the adversarial suffix of large language model attacks via translation of text embeddings.CoRR2024.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models From noise to clarity: Unraveling the adversarial suffix of large language model attacks via translation of text embeddings.CoRR2024

Reference 42

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T23:21:01.748411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:56.742083Z digest=sha256:b6412eefbdfa10c4bb399d0a7ae5110b7973c57a931801f4c90c7c93f4600c64

Observation 0f629e3d-fe27-4fae-93c7-e2ab5756f978 · outbound

This paper cites Guiding not Forcing: Enhancing the Transferability of Jailbreaking Attacks on LLMs via Removing Superfluous Constraints.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Guiding not Forcing: Enhancing the Transferability of Jailbreaking Attacks on LLMs via Removing Superfluous Constraints

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:21:00.004770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:56.844753Z digest=sha256:2ce11e3e8c45562057aa9a77e3dff60fa28b10cd7b2b1ca4ee036cd990c63a9d

Observation 73d991a7-641e-443f-9e8e-375b9bf881d5 · outbound

This paper cites AutoDAN: Interpretable Gradient-Based Adversarial Attacks on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models AutoDAN: Interpretable Gradient-Based Adversarial Attacks on Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:56.924751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:56.924751Z digest=sha256:8f6d88ad5ae5c677079b4a0b86794facd93058618327b554d67d8e9f1dc9c64b

Observation f2de84f1-870e-469e-b6f7-00b81abe62f6 · outbound

This paper cites Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.014753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.014753Z digest=sha256:a0618ba093079529477633ec6c3fd1bd9964370a0438a5f90bca1563f678ee3a

Observation d6c92b03-eb4c-412c-a676-26de960e4f10 · outbound

This paper cites COLD-Attack: Jailbreaking LLMs with Stealthiness and Controllability.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models COLD-Attack: Jailbreaking LLMs with Stealthiness and Controllability

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:01.592880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:57.106875Z digest=sha256:e44c15b40129ff826fbef795d4e452bfb33ad0043eef0a5e1032d8fb09b59869

Observation 4340b1d7-49c0-436b-84dc-94c912fa73fd · outbound

This paper cites DROJ: A Prompt-Driven Attack against Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models DROJ: A Prompt-Driven Attack against Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.184523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.184523Z digest=sha256:0e79c584b37c3a4bf71b1c5217c9a22457f4655f4cc70d1cee5a99196cccd338

Observation 31f963ff-fdbe-4b35-b7e8-2b9f30b76df0 · outbound

This paper cites Catastrophic jailbreak of open-source llms via exploiting generation.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Catastrophic jailbreak of open-source llms via exploiting generation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:01.378863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:57.318459Z digest=sha256:1ebc1daf8f40a22f6eb551fea90e45c98fa2cff966ea38340722d870ab30143c

Observation dbf5d646-15db-44cb-995c-4f51ac048dd8 · outbound

This paper cites Make Them Spill the Beans! Coercive Knowledge Extraction from (Production) LLMs.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Make Them Spill the Beans! Coercive Knowledge Extraction from (Production) LLMs

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.420571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.420571Z digest=sha256:dbce4af3cfd4c0992219e756d40c40c102ab06796a2c92157f9e9a33299786a8

Observation 31368c05-277e-4bbf-b0e3-6e2c92dff6f8 · outbound

This paper cites Weak-to-Strong Jailbreaking on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Weak-to-Strong Jailbreaking on Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.504508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.504508Z digest=sha256:7e2cc30afa6e76df8d9e4d25cb27a9c3b9b51871519fb6c731e5b13c86f4f46c

Observation 42a48662-4628-4735-83ee-9602c01ffcd3 · outbound

This paper cites Don't Say No: Jailbreaking LLM by Suppressing Refusal.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.656839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.656839Z digest=sha256:4ea71d89bf2e642918bc95907e35822354534ca54695c955256570a3aadb8ea4

Observation 42f69fb9-0057-4e6d-88e6-42a3df17cf2d · outbound

This paper cites Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.744842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.744842Z digest=sha256:25882458f8969fe1df1190e7dc24e3af37035d74da55ba8fff3531e19c71c7cb

Observation 8423a4ca-4720-4e34-9c9a-ca8ba783decd · outbound

This paper cites Shadow Alignment: The Ease of Subverting Safely-Aligned Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Shadow Alignment: The Ease of Subverting Safely-Aligned Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.826882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.826882Z digest=sha256:a38f0b02bd30cbe859abbef2c4ba66884dc2f682d56ed66b7e1f3d9cb9becd81

Observation ab1d7d35-a32d-4fb2-9219-90097f0a8d86 · outbound

This paper cites Removing RLHF Protections in GPT-4 via Fine-Tuning.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Removing RLHF Protections in GPT-4 via Fine-Tuning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.988075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.988075Z digest=sha256:a17f80359c8a33b9c9319143c3462d2301349f360469d0e1a6327ba20b913140

Observation 685c4854-50a5-45aa-9108-651a18eea961 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.119677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.119677Z digest=sha256:26bbfdb28443addabdeae9a7a14aaf1a69aebef8c6a2b61c25d3cd6e29a58205

Observation 8ace8b2b-0596-49af-b719-c484030b59f7 · outbound

This paper cites EasyJailbreak: A Unified Framework for Jailbreaking Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models EasyJailbreak: A Unified Framework for Jailbreaking Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.217814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.217814Z digest=sha256:297c388ce726b2959c2d2e8468b0d2f9971506905324640ef429f86af08ae6e0

Observation d8e38e2c-5701-411c-be1a-3df3c7fb0434 · outbound

This paper cites Playing Language Game with LLMs Leads to Jailbreaking.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Playing Language Game with LLMs Leads to Jailbreaking

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.265798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.265798Z digest=sha256:2e2bca29180d3d6201ef44d9982691ea01facb297cc7d1fad961affeda7c9af1

Observation 8453da1f-f303-4dfc-a528-6b3a81731ada · outbound

This paper cites do anything now.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models do anything now

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:01.234729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:58.301037Z digest=sha256:5ce8000924181c91a5a14cab550281224b563baeb478e903b61c1825266568d6

Observation 9d70e8b5-b0ee-4355-816b-25171df3f631 · outbound

This paper cites Chain-of-Lure: A Synthetic Narrative-Driven Approach to Compromise Large Language Models.arXiv preprint arXiv:2505.175192025.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Chain-of-Lure: A Synthetic Narrative-Driven Approach to Compromise Large Language Models.arXiv preprint arXiv:2505.175192025

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.396202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.396202Z digest=sha256:9d1b9da549e365b1e8fa4c8f02fa64f0858d3ee50e25ead168ab8264dcfe69fc

Observation 9cfac2d0-946a-46f1-b1ce-98acfb8deaf1 · outbound

This paper cites Reasoning-Augmented Conversation for Multi-Turn Jailbreak Attacks on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Reasoning-Augmented Conversation for Multi-Turn Jailbreak Attacks on Large Language Models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.435774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.435774Z digest=sha256:3fe0209ca83cde05872c7ea88ee55d82115dfe9526efaecb54e55a591018b33e

Observation 12e3a4ea-66ab-4847-8f75-b609b11b7db2 · outbound

This paper cites Amplified Vulnerabilities: Structured Jailbreak Attacks on LLM-based Multi-Agent Debate.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Amplified Vulnerabilities: Structured Jailbreak Attacks on LLM-based Multi-Agent Debate

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.553566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.553566Z digest=sha256:570c76f1099a095a44af2e4baed45f55025b56e0c8f62baea46c097e510f37a5

Observation a5edf243-33c2-4ba1-a47c-b019af836f53 · outbound

This paper cites H-CoT: Hijacking the Chain-of-Thought Safety Reasoning Mechanism to Jailbreak Large Reasoning Models, Including OpenAI o1/o3, DeepSeek-R1, and Gemini 2.0 Flash Thinking.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models H-CoT: Hijacking the Chain-of-Thought Safety Reasoning Mechanism to Jailbreak Large Reasoning Models, Including OpenAI o1/o3, DeepSeek-R1, and Gemini 2.0 Flash Thinking

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.633864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.633864Z digest=sha256:04e9dabe05f531c3d97115b98192e11633b15a4c7085f80b02753a9509f8544b

Observation 5da767d8-d948-43cd-ae62-0c59201e73a0 · outbound

This paper cites Advancing Jailbreak Strategies: A Hybrid Approach to Exploiting LLM Vulnerabilities and Bypassing Modern Defenses.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Advancing Jailbreak Strategies: A Hybrid Approach to Exploiting LLM Vulnerabilities and Bypassing Modern Defenses

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.767251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.767251Z digest=sha256:0baaf9772f09d555d17707ffe114ff1f3deec56c71b65763257406e988c07677

Observation 780df3e5-5c8b-4e64-9949-794827c298c4 · outbound

This paper cites Gradient cuff: Detecting jailbreak attacks on large language models by exploring refusal loss landscapes.NeurIPS2024,37, 126265–126296.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Gradient cuff: Detecting jailbreak attacks on large language models by exploring refusal loss landscapes.NeurIPS2024,37, 126265–126296

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:01.066495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:58.874995Z digest=sha256:ece1755af354f690f844fed7187e364ee4255f52b7312ebb9a7045263f9ad901

Observation 4b26e0de-d342-4c23-b32a-554f30637313 · outbound

This paper cites JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:00.942413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:58.957828Z digest=sha256:c3e80e1273093e1cc4049f676257ff5865c3905d6c39542434b43a5345fdd74d

Observation acc9b9a0-018d-4138-8dfa-666ea874745b · outbound

This paper cites The hidden risks of large reasoning models: A safety assessment of r1.arXiv preprint arXiv:2502.126592025.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models The hidden risks of large reasoning models: A safety assessment of r1.arXiv preprint arXiv:2502.126592025

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:59.040436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:59.040436Z digest=sha256:956c45581580031eec4241969a7a8d83af51d550fe70184226c649de0b2f93ae

Observation 8625a784-0c3b-4c2e-903f-810fca884fe2 · outbound

This paper cites Scaling Trends in Language Model Robustness.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Scaling Trends in Language Model Robustness

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:00.662149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:59.108858Z digest=sha256:6bf1989b94efdd5b627346167d0223d55106a92d04075caf251e953da367cfb4

Observation 9d2998c7-079e-4b37-8077-327cbdf5afaf · outbound

This paper cites Scaling Behavior of Machine Translation with Large Language Models under Prompt Injection Attacks.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Scaling Behavior of Machine Translation with Large Language Models under Prompt Injection Attacks

Reference 68

Resolution
malformed identifier
local_arxiv, observed 2026-08-06T23:20:59.349100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T23:20:59.199192Z digest=sha256:09deb3d4b6693dfa30e7f4c230a158fa70b2811deb8dff56196c6073977940b3

Pith citing papers

No inbound Pith citation observations are available.