Pith. sign in

Paper Citation Record · LEDGER

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models

As of 22 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 0 inbound Pith citation observations for arXiv:2506.18543.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.18543 v2

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:20:59.199192Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

68 of 68 outbound references displayed

  • verified exact1
  • verified fuzzy22
  • unresolved42
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation aed38e89-35de-4a6d-921c-1a46b2b99efa · outbound

This paper cites Improving language understanding by generative pre-training2018.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Improving language understanding by generative pre-training2018

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:04.904571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:51.740664Z digest=sha256:3b0eb763321069f02a412011dd87ba774c1fb0911173b3d65f9e5942d84a7726

Observation 1ef3f846-4409-4170-a3c4-4c062658fd0d · outbound

This paper cites GPT-4 Technical Report.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:51.894501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:51.894501Z digest=sha256:9a75bec1f79c491fa25f1dcf33daf2c311aceb6899bc91b373db43c5ae4ee6f9

Observation d3dd119b-d654-4d6b-b168-1bdb893da6c2 · outbound

This paper cites DeepSeek LLM: Scaling Open-Source Language Models with Longtermism.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:51.997871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:51.997871Z digest=sha256:bf226f08aba94aba9418b2a7e2bc0431566053c3ecda9dcaad18fb269fe7b807

Observation 475ed272-df6e-433f-8369-58ebefe17c1a · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:52.145353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:52.145353Z digest=sha256:71a0c497726efd96f0161670a4254defe153b7371e163443165e6d6439858df4

Observation 2f51017b-8333-451b-9117-773d2e2a4d40 · outbound

This paper cites PAL: Proxy-Guided Black-Box Attack on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:52.222527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:52.222527Z digest=sha256:a899407f48e05c0f8cf56658f16c894aebc33487e052fb58468af377e5713495

Observation 397bd4f2-0c8f-40b2-bb8e-1e50f9542ab5 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:52.358358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:52.358358Z digest=sha256:beb8f9c5e34b12ff3d6da7d2cbe3f54e7f624669f0551b233b01100d92603407

Observation cc1567fc-3ba2-4fc9-9c42-1ccde88ab3d8 · outbound

This paper cites Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:52.542586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:52.542586Z digest=sha256:d3bca4a5f7f0bf3e98a580aab176eb84da531e4eaaf989eaa7b1ce4cd21e6d26

Observation 3cf6324c-550c-4d8a-8cbb-105e3cb79ade · outbound

This paper cites A survey of backdoor attacks and defenses on large language models: Implications for security measures.Authorea Preprints2024.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models A survey of backdoor attacks and defenses on large language models: Implications for security measures.Authorea Preprints2024

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:04.636289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:52.790543Z digest=sha256:f9b1222cc7ad9c63286a383e959c2bfbc2ed970fa75a5d7e4ccf34c6eef3cd84

Observation b8189ea6-8014-43c2-9a13-f58b06787a87 · outbound

This paper cites Training language models to follow instructions with human feedback.NeurIPS2022,35, 27730–27744.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Training language models to follow instructions with human feedback.NeurIPS2022,35, 27730–27744

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:04.274233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:52.876347Z digest=sha256:975a99f2e4c2e3236c3e1eda1373b6af66f5b839b8b1d0f493298120b9ea7cea

Observation c6aa9a52-c618-443e-b245-10472b891979 · outbound

This paper cites Defending Large Language Models Against Jailbreak Attacks Through Chain of Thought Prompting.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Defending Large Language Models Against Jailbreak Attacks Through Chain of Thought Prompting

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:04.119138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:53.014644Z digest=sha256:12148cb5ff9d3585fce0787da63d797b316cfa5bb43d6a156ffd635b66459b28

Observation ad0c9440-6356-4961-8102-c23bd7bf22f1 · outbound

This paper cites Many-shot jailbreaking.NeurIPS2024,37, 129696–129742.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Many-shot jailbreaking.NeurIPS2024,37, 129696–129742

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:03.956256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:53.104384Z digest=sha256:8ed2b1a96c8987a4873caaf8ecfeab25aa7aac9fe026088209178223388e1920

Observation 1107255d-f51b-401a-841e-7976b9a6d564 · outbound

This paper cites Pandora: Jailbreak GPTs by Retrieval Augmented Generation Poisoning.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Pandora: Jailbreak GPTs by Retrieval Augmented Generation Poisoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:53.234749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:53.234749Z digest=sha256:ef0a16b54540f29a059d812df0e8aa8cfd306cc5440392dac3071a38729bd512

Observation 2ec64b4f-08f6-42c6-a595-c514a2f21e81 · outbound

This paper cites Multi-step Jailbreaking Privacy Attacks on ChatGPT.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Multi-step Jailbreaking Privacy Attacks on ChatGPT

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:53.545039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:53.545039Z digest=sha256:054e301e637d1dc55a7edd2efb1a5c89e686dc7ddde9cb5ff971c7a6bd01c75b

Observation 5b09016f-1598-49f5-8296-9045b92df4b2 · outbound

This paper cites Adversarial Demonstration Attacks on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Adversarial Demonstration Attacks on Large Language Models

Reference 14

Resolution
malformed identifier
no resolver link, observed 2026-08-06T23:20:53.674318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:53.674318Z digest=sha256:a05c650efa895c9c546580de01f0b87d6b8d251913d86f901e2a757e12bb502f

Observation acd11a48-ed3f-4c4d-be94-32473105ddd4 · outbound

This paper cites Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:53.797609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:53.797609Z digest=sha256:e0c5d8d8539f3ef1ddd3db1ed6ad97f35956c9c1a37a1ac6aa42f5845baa2d3b

Observation 8c70371a-3c8c-4c98-8d98-02c6c726e70b · outbound

This paper cites Improved few-shot jailbreaking can circumvent aligned language models and their defenses.NeurIPS2024,37, 32856–32887.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Improved few-shot jailbreaking can circumvent aligned language models and their defenses.NeurIPS2024,37, 32856–32887

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:03.744360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:53.856016Z digest=sha256:9a4ce5779ff962797d5e7a23dbac3595c7f3a5e17d81b64dea5315aa859ed05f

Observation c19d5bb5-90b4-41b0-a6fe-a093ffc61213 · outbound

This paper cites Play Guessing Game with LLM: Indirect Jailbreak Attack with Implicit Clues.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Play Guessing Game with LLM: Indirect Jailbreak Attack with Implicit Clues

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:53.974751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:53.974751Z digest=sha256:95c8e2fdb1bca1b146639ac86a27da3146e4475a35f851734f460352592427eb

Observation abd74381-d8d7-4181-8228-b93e0156df97 · outbound

This paper cites Artprompt: Ascii art-based jailbreak attacks against aligned llms.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Artprompt: Ascii art-based jailbreak attacks against aligned llms

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:03.585017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:54.051307Z digest=sha256:03e7b714bb2b2c54582f8926c51ded9993e1b81544a14ef7abe4607098fee5a2

Observation 7934abf5-bb74-4271-a43f-dfa0858bc75c · outbound

This paper cites Semantic Mirror Jailbreak: Genetic Algorithm Based Jailbreak Prompts Against Open-source LLMs.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Semantic Mirror Jailbreak: Genetic Algorithm Based Jailbreak Prompts Against Open-source LLMs

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:54.194806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:54.194806Z digest=sha256:1e581050faed05cd5469deb5b8f3b464c67ca7fa12e9cc614efc9489312ac672

Observation 62c6b020-c4b5-4cfb-ba2b-4431d76ecd1d · outbound

This paper cites Understanding and Enhancing the Transferability of Jailbreaking Attacks.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Understanding and Enhancing the Transferability of Jailbreaking Attacks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:54.264785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:54.264785Z digest=sha256:2e1747a768141e3e31e0dd36015c10d0f5f47c82979e53c4ddf228dcc9047cb3

Observation 04bf703d-e726-4859-8139-7bd43ff41a54 · outbound

This paper cites AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:54.385546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:54.385546Z digest=sha256:9a8bb91aaf3a453041541533072929d0e6399b6dc167635001a0e61a97a67d5e

Observation f737caa1-dd19-4063-b49f-336cf5337eaf · outbound

This paper cites Flipattack: Jailbreak llms via flipping.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Flipattack: Jailbreak llms via flipping

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:03.365496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:54.467206Z digest=sha256:7765d581aa986c67dc1a8720a07db8aae191955ac8bec079595a35ab92b9432a

Observation b80279ff-90d9-451b-bac6-6d4c9d7bfbff · outbound

This paper cites All in how you ask for it: Simple black-box method for jailbreak attacks.Applied Sciences2024,14, 3558.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models All in how you ask for it: Simple black-box method for jailbreak attacks.Applied Sciences2024,14, 3558

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:03.130334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:54.614751Z digest=sha256:84a6a969d79f5569d7b3b4488cc451705b8dc2a5c16f6f63b025f98cd45c43ab

Observation 66031097-22d5-4869-95db-3bd8274d5547 · outbound

This paper cites Emoji Attack: Enhancing Jailbreak Attacks Against Judge LLM Detection.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Emoji Attack: Enhancing Jailbreak Attacks Against Judge LLM Detection

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:02.924635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:54.718877Z digest=sha256:0a090f58cd3f14e33337e52a697b0f60bb4146dee063f72ebe29c533d96044c2

Observation 572595dc-71e9-4f77-ab83-eafeb7994cbd · outbound

This paper cites The Dark Side of Trust: Authority Citation-Driven Jailbreak Attacks on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models The Dark Side of Trust: Authority Citation-Driven Jailbreak Attacks on Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:54.845285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:54.845285Z digest=sha256:674ce51ae9bbbea85046db33bcef88e26202c0f23f18324cb1cddc5e72b33dbc

Observation 80a5b13b-f9b0-4d0d-8bd2-195e885fbab4 · outbound

This paper cites GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:54.935278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:54.935278Z digest=sha256:63f5d6bfb6aba177dd6966bf4ca70a9e4e750bb1ed329b7cebad461d87c257d6

Observation fe8b4cf5-c9ed-4ddc-9b02-460cffe25f72 · outbound

This paper cites Gpt-4 is too smart to be safe: Stealthy chat with llms via cipher2024.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Gpt-4 is too smart to be safe: Stealthy chat with llms via cipher2024

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:02.700975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:55.064753Z digest=sha256:f49d0b466246fa835d0fefee32fcc5ddb7914479913a920c0b5a5e3f7814eca7

Observation 54a66195-b915-43e4-9959-84fd95786d5c · outbound

This paper cites Explore, Establish, Exploit: Red Teaming Language Models from Scratch.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Explore, Establish, Exploit: Red Teaming Language Models from Scratch

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:55.144189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:55.144189Z digest=sha256:7c874ea3a5f10953df46b5e3dc76735878899d2ac2617c7310acd9c59a5939c5

Observation c04a6230-7c1e-48d7-a277-2891647b17b5 · outbound

This paper cites Jailbreaking black box large language models in twenty queries.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Jailbreaking black box large language models in twenty queries

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:02.554468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:55.255225Z digest=sha256:8da2e60414a16dbe76fa88039c8125c3b9b4c65922d9d24a072d2e62d445e213

Observation 4e48964a-e6cf-40b9-95be-f7f0537b1c4b · outbound

This paper cites MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:55.350714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:55.350714Z digest=sha256:b573acf621dc2f3850f04cc85a82760a426e2d572f3618b8c5522b49e1a06f23

Observation 55094c79-bba4-4d98-b270-58aefc9fa2a9 · outbound

This paper cites MART: Improving LLM Safety with Multi-round Automatic Red-Teaming.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models MART: Improving LLM Safety with Multi-round Automatic Red-Teaming

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:02.386229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:55.425537Z digest=sha256:e66a8905af1fea61ecff0bfb000367a8818f746b771ec72263863c44a758195e

Observation 4c262b5e-e8cb-4176-8743-7c89d623960d · outbound

This paper cites Guard: Role-playing to generate natural-language jailbreakings to test guideline adherence of large language models.arXiv preprint arXiv:2402.032992024.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Guard: Role-playing to generate natural-language jailbreakings to test guideline adherence of large language models.arXiv preprint arXiv:2402.032992024

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:55.484326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:55.484326Z digest=sha256:7cf7e72a78a69150535186aab3d21b24d205c543158421fe5c12d87d99a0a847

Observation 1fc88740-52ae-418d-8b91-0abd30c85890 · outbound

This paper cites Goal-Oriented Prompt Attack and Safety Evaluation for LLMs.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Goal-Oriented Prompt Attack and Safety Evaluation for LLMs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:55.602649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:55.602649Z digest=sha256:93ce04aabd424df7ef794421b67c1f7c7df73918f4772dc7838f7911c03adfd9

Observation 4ed0e8b6-e09e-46ab-8341-c7523c28610e · outbound

This paper cites Tree of attacks: Jailbreaking black-box llms automatically.NeurIPS2024,37, 61065–61105.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Tree of attacks: Jailbreaking black-box llms automatically.NeurIPS2024,37, 61065–61105

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:02.212017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:55.752634Z digest=sha256:38c662ce6ed75368be5278a314ed4e709185a5b0f6c182a8102f2a63aaef887f

Observation 6d906230-1a4e-4b3f-b9a8-854bff233fbf · outbound

This paper cites Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Scalable and Transferable Black-Box Jailbreaks for Language Models via Persona Modulation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:55.859123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:55.859123Z digest=sha256:6bf7b91e43e0a2ff69754b6ef4723b58ed810764fc54cd858979a6f6f92ff144

Observation 207ab80e-bb35-4fe7-b1ba-5e6a8bd3c6af · outbound

This paper cites Evil Geniuses: Delving into the Safety of LLM-based Agents.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Evil Geniuses: Delving into the Safety of LLM-based Agents

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:55.953808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:55.953808Z digest=sha256:25b96794a49d0cd01ce18171453983f7639dab6265d9aa19a7bea60c88ba5c5e

Observation 898e97be-5c05-4f88-9ce5-c472c57af4b6 · outbound

This paper cites How johnny can persuade llms to jailbreak them: Rethinking persuasion to challenge ai safety by humanizing llms.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models How johnny can persuade llms to jailbreak them: Rethinking persuasion to challenge ai safety by humanizing llms

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:02.015166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:56.055295Z digest=sha256:47eb42862789923cfcfa026526a0810791a2d90c8e536b2a2ea7b85615e563c8

Observation 09cebd2c-7789-45e8-89f2-5eb8bf3a0b10 · outbound

This paper cites Attacking Large Language Models with Projected Gradient Descent.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Attacking Large Language Models with Projected Gradient Descent

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:56.204745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:56.204745Z digest=sha256:d2468242eea70e55b93be16c15ca634c848862f73568a6f3ee0d3bdd240cd3f9

Observation a469243d-98ab-448f-a553-d9a672b4148f · outbound

This paper cites Query-based adversarial prompt generation.NeurIPS2024, 37, 128260–128279.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Query-based adversarial prompt generation.NeurIPS2024, 37, 128260–128279

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:01.842246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:56.319414Z digest=sha256:ac1c1d841eb7c962ba5a5bb0b06ff71beb113b70fcf9b35e29bb2a525ff9f5c1

Observation d2bfecdd-1356-4824-8494-fac5065e7174 · outbound

This paper cites Improved Techniques for Optimization-Based Jailbreaking on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Improved Techniques for Optimization-Based Jailbreaking on Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:56.482120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:56.482120Z digest=sha256:065b3c69a26de658bf92b670f59e6e9bfa883df90b7a24261cbc119b5f559830

Observation 2309b2d9-7282-4885-9774-dc260c466e86 · outbound

This paper cites Iterative self-tuning llms for enhanced jailbreaking capabilities.arXiv preprint arXiv:2410.184692024.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Iterative self-tuning llms for enhanced jailbreaking capabilities.arXiv preprint arXiv:2410.184692024

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:56.584927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:56.584927Z digest=sha256:05c1597e5c909cf3382251f75f8e562a60f3de336a2ebb4f6957052dddba1e57

Observation 980df864-1235-45f1-a729-c8df39e53787 · outbound

This paper cites From noise to clarity: Unraveling the adversarial suffix of large language model attacks via translation of text embeddings.CoRR2024.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models From noise to clarity: Unraveling the adversarial suffix of large language model attacks via translation of text embeddings.CoRR2024

Reference 42

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T23:21:01.748411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:56.742083Z digest=sha256:7125379635e951891b2543c1f437b7cb40e8b667d7ea97306355fcc0b92789a6

Observation 0f629e3d-fe27-4fae-93c7-e2ab5756f978 · outbound

This paper cites Guiding not Forcing: Enhancing the Transferability of Jailbreaking Attacks on LLMs via Removing Superfluous Constraints.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Guiding not Forcing: Enhancing the Transferability of Jailbreaking Attacks on LLMs via Removing Superfluous Constraints

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:21:00.004770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:56.844753Z digest=sha256:a5347bc50fd87f88a8c952b5a24e25142bec9f899ec812ebf33113416d58ae43

Observation 73d991a7-641e-443f-9e8e-375b9bf881d5 · outbound

This paper cites AutoDAN: Interpretable Gradient-Based Adversarial Attacks on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models AutoDAN: Interpretable Gradient-Based Adversarial Attacks on Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:56.924751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:56.924751Z digest=sha256:51e21d8894e7725e0b0382e215a7c60e798e2230737e8db654eee9f73757fc61

Observation f2de84f1-870e-469e-b6f7-00b81abe62f6 · outbound

This paper cites Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Analyzing the Inherent Response Tendency of LLMs: Real-World Instructions-Driven Jailbreak

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.014753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.014753Z digest=sha256:e6d7f8587214746cf40adffe0d9b6edea8178e44779adca4b040ee3777de3b7f

Observation d6c92b03-eb4c-412c-a676-26de960e4f10 · outbound

This paper cites COLD-Attack: Jailbreaking LLMs with Stealthiness and Controllability.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models COLD-Attack: Jailbreaking LLMs with Stealthiness and Controllability

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:01.592880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:57.106875Z digest=sha256:bc92ffecd5f0202d6053b2f8c15f6aa1de52cbc740c2df2fff0fd46a64cc28ca

Observation 4340b1d7-49c0-436b-84dc-94c912fa73fd · outbound

This paper cites DROJ: A Prompt-Driven Attack against Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models DROJ: A Prompt-Driven Attack against Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.184523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.184523Z digest=sha256:998cc79a67f55680a028113c0efd28a8746e852274275318683f4719394bf458

Observation 31f963ff-fdbe-4b35-b7e8-2b9f30b76df0 · outbound

This paper cites Catastrophic jailbreak of open-source llms via exploiting generation.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Catastrophic jailbreak of open-source llms via exploiting generation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:01.378863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:57.318459Z digest=sha256:cf9ff08bbacf472d37420059757664b7cd79960e01c9cf2bc07295fb90353341

Observation dbf5d646-15db-44cb-995c-4f51ac048dd8 · outbound

This paper cites Make Them Spill the Beans! Coercive Knowledge Extraction from (Production) LLMs.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Make Them Spill the Beans! Coercive Knowledge Extraction from (Production) LLMs

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.420571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.420571Z digest=sha256:e380a589d7b64f05da57ffeb85d2b2af8b8fe3a9fc7d0bf9cef267586c136e11

Observation 31368c05-277e-4bbf-b0e3-6e2c92dff6f8 · outbound

This paper cites Weak-to-Strong Jailbreaking on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Weak-to-Strong Jailbreaking on Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.504508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.504508Z digest=sha256:6750dd32424b84030e6d0daf8741a1cdd46a1f7d9e4cc2abd0023c470f901144

Observation 42a48662-4628-4735-83ee-9602c01ffcd3 · outbound

This paper cites Don't Say No: Jailbreaking LLM by Suppressing Refusal.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.656839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.656839Z digest=sha256:3e9a003a659b6320511f7f03db9a7e0a0c4e407c4c265a32f622bfd44a1a820a

Observation 42f69fb9-0057-4e6d-88e6-42a3df17cf2d · outbound

This paper cites Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.744842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.744842Z digest=sha256:22020264b09ccb1902dffd42af06ac968c08aca753f47b3aee434f4bd7c8c888

Observation 8423a4ca-4720-4e34-9c9a-ca8ba783decd · outbound

This paper cites Shadow Alignment: The Ease of Subverting Safely-Aligned Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Shadow Alignment: The Ease of Subverting Safely-Aligned Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.826882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.826882Z digest=sha256:14a091b442acce0af327bfc75ca3585ae77d8d83414b7700f7dda5ae73b92283

Observation ab1d7d35-a32d-4fb2-9219-90097f0a8d86 · outbound

This paper cites Removing RLHF Protections in GPT-4 via Fine-Tuning.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Removing RLHF Protections in GPT-4 via Fine-Tuning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.988075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.988075Z digest=sha256:5c42726b0b0df73a1d1368bbcadffd18d1faeb19efb20fada317b000bc3e9295

Observation 685c4854-50a5-45aa-9108-651a18eea961 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.119677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.119677Z digest=sha256:fe4cc85ac7be09e1d24f0dd894d50d061396b94814c5a112d9b9cca5e23ad8ab

Observation 8ace8b2b-0596-49af-b719-c484030b59f7 · outbound

This paper cites EasyJailbreak: A Unified Framework for Jailbreaking Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models EasyJailbreak: A Unified Framework for Jailbreaking Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.217814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.217814Z digest=sha256:2a5c4b26ccb5bdedf3c4f8d3c2b399873da142b2bac7fbde5fab0eab7be29aa4

Observation d8e38e2c-5701-411c-be1a-3df3c7fb0434 · outbound

This paper cites Playing Language Game with LLMs Leads to Jailbreaking.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Playing Language Game with LLMs Leads to Jailbreaking

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.265798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.265798Z digest=sha256:f92bddc3f56109209788d28d1d6ce80ddb043f92dd4b43b690794e7c3bed2b80

Observation 8453da1f-f303-4dfc-a528-6b3a81731ada · outbound

This paper cites do anything now.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models do anything now

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:01.234729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:58.301037Z digest=sha256:d13ffd1fe900c07eea89b8039a64b4484aec3a30b3a8ba16093a37a9e093051a

Observation 9d70e8b5-b0ee-4355-816b-25171df3f631 · outbound

This paper cites Chain-of-Lure: A Synthetic Narrative-Driven Approach to Compromise Large Language Models.arXiv preprint arXiv:2505.175192025.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Chain-of-Lure: A Synthetic Narrative-Driven Approach to Compromise Large Language Models.arXiv preprint arXiv:2505.175192025

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.396202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.396202Z digest=sha256:c6af9ae6da078c7bef448f2abf205c375e0c713652d360ec2432e0fdb7f730b5

Observation 9cfac2d0-946a-46f1-b1ce-98acfb8deaf1 · outbound

This paper cites Reasoning-Augmented Conversation for Multi-Turn Jailbreak Attacks on Large Language Models.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Reasoning-Augmented Conversation for Multi-Turn Jailbreak Attacks on Large Language Models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.435774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.435774Z digest=sha256:25cb2c30932c911a7d042040a802f19c81858a58eb47b74aef1f46544e51f39b

Observation 12e3a4ea-66ab-4847-8f75-b609b11b7db2 · outbound

This paper cites Amplified Vulnerabilities: Structured Jailbreak Attacks on LLM-based Multi-Agent Debate.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Amplified Vulnerabilities: Structured Jailbreak Attacks on LLM-based Multi-Agent Debate

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.553566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.553566Z digest=sha256:88dc2f58a4bd66ba5daefe9596a7dece814d9819a7132b868f301b1d48e8223a

Observation a5edf243-33c2-4ba1-a47c-b019af836f53 · outbound

This paper cites H-CoT: Hijacking the Chain-of-Thought Safety Reasoning Mechanism to Jailbreak Large Reasoning Models, Including OpenAI o1/o3, DeepSeek-R1, and Gemini 2.0 Flash Thinking.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models H-CoT: Hijacking the Chain-of-Thought Safety Reasoning Mechanism to Jailbreak Large Reasoning Models, Including OpenAI o1/o3, DeepSeek-R1, and Gemini 2.0 Flash Thinking

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.633864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.633864Z digest=sha256:18715f11c08e095b70f9cc83bfea4f12c05b9dbb82ed1f6db3b6caa53dee8f24

Observation 5da767d8-d948-43cd-ae62-0c59201e73a0 · outbound

This paper cites Advancing Jailbreak Strategies: A Hybrid Approach to Exploiting LLM Vulnerabilities and Bypassing Modern Defenses.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Advancing Jailbreak Strategies: A Hybrid Approach to Exploiting LLM Vulnerabilities and Bypassing Modern Defenses

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:58.767251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:58.767251Z digest=sha256:865fe521d78a7d7dd2b22c84ff651909c834654b0792aa4d23c4d04127844283

Observation 780df3e5-5c8b-4e64-9949-794827c298c4 · outbound

This paper cites Gradient cuff: Detecting jailbreak attacks on large language models by exploring refusal loss landscapes.NeurIPS2024,37, 126265–126296.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Gradient cuff: Detecting jailbreak attacks on large language models by exploring refusal loss landscapes.NeurIPS2024,37, 126265–126296

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:01.066495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:58.874995Z digest=sha256:7088779d7f3b9ea33806467c686a6788464ea348e30bc6d0559194e216d33868

Observation 4b26e0de-d342-4c23-b32a-554f30637313 · outbound

This paper cites JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:00.942413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:58.957828Z digest=sha256:6c9de36924c73775b4b01e9bcaafe8b6d9a588e8f68a74733bdb53fac2226c11

Observation acc9b9a0-018d-4138-8dfa-666ea874745b · outbound

This paper cites The hidden risks of large reasoning models: A safety assessment of r1.arXiv preprint arXiv:2502.126592025.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models The hidden risks of large reasoning models: A safety assessment of r1.arXiv preprint arXiv:2502.126592025

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:59.040436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:59.040436Z digest=sha256:f3da5840074b67efe18fa4990386ee19917a6e686d321711eeed157960bd5141

Observation 8625a784-0c3b-4c2e-903f-810fca884fe2 · outbound

This paper cites Scaling Trends in Language Model Robustness.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Scaling Trends in Language Model Robustness

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:21:00.662149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:59.108858Z digest=sha256:b1d105043af9ddbc95ce0917680e15fc2d1234d87e079d495a0d617a2ea7b67a

Observation 9d2998c7-079e-4b37-8077-327cbdf5afaf · outbound

This paper cites Scaling Behavior of Machine Translation with Large Language Models under Prompt Injection Attacks.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Scaling Behavior of Machine Translation with Large Language Models under Prompt Injection Attacks

Reference 68

Resolution
malformed identifier
local_arxiv, observed 2026-08-06T23:20:59.349100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:20:59.199192Z digest=sha256:97d20d24209c3829fb419c77d2910b9bd22ab428386db0e96afb17368e3511ff

Pith citing papers

No inbound Pith citation observations are available.