Pith. sign in

Paper Citation Record · LEDGER

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts

As of 8 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 2 inbound Pith citation observations for arXiv:2506.07596.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07596 v1

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:36:58.179549Z

measured 77 of 77 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T01:03:28.012736Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

75 of 75 outbound references displayed

  • verified exact0
  • verified fuzzy42
  • unresolved30
  • parse uncertain1
  • malformed identifier2
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation c4667882-bb7d-4cfd-85df-ff0099d9668e · outbound

This paper cites DeepSeek LLM 7B Chat.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts DeepSeek LLM 7B Chat

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.375949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.811450Z digest=sha256:cbd6735b8a32925b42f3620cbb5c98afcae31c89a6f56c251ce16abfd63f196d

Observation 3d7b51bc-6d31-47f2-8e95-ad80061bfea1 · outbound

This paper cites Mistral 7B Instruct v0.2.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Mistral 7B Instruct v0.2

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.363168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.817870Z digest=sha256:3bb102ed528ab286b41f2193e1843a5d58ad886ae69ab6e3972d0a54632dc032

Observation c069bfaa-ee30-474d-a255-197e21c74855 · outbound

This paper cites Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:57.822710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:57.822710Z digest=sha256:4f9a78b48642d8dee5f5823949f1b1b37a2c71936be146fe0fda54396efa2c34

Observation 1acdfa2e-b844-4d38-9089-dd432d6eead2 · outbound

This paper cites an unresolved cited work.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:36:59.349778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.829422Z digest=sha256:86405f6196965075805f3057835ca0fb1e57b61c7caf99a3d257f366d6df643e

Observation f2d19aa2-596e-4e64-b8d3-bfcfd339d83e · outbound

This paper cites Language Models are Few-Shot Learn- ers.NeurIPS, 2020.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Language Models are Few-Shot Learn- ers.NeurIPS, 2020

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.336081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.834491Z digest=sha256:0b4f47185f6bd354e0f56017560ee86c0d045bf21759be4108708532223c08b1

Observation a331f1c4-8f63-4135-9629-dcc32f00318b · outbound

This paper cites A review of the application of deep learning in medical image classifica- tion and segmentation.Annals of translational medicine, 2020.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts A review of the application of deep learning in medical image classifica- tion and segmentation.Annals of translational medicine, 2020

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.322660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.840367Z digest=sha256:349b3b6dd9fa824ee9ed9fcb2ac4ecdcbda7fba4639d1d95988fe9eec8e67251

Observation 0655e8a1-bcb6-4e6a-8913-731f5ce195af · outbound

This paper cites Pappas, and Eric Wong.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Pappas, and Eric Wong

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.307761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.845714Z digest=sha256:bab4b120541dc37c73d0e47ebc1f81ee10a8af6b439f3c5c50cfe11f8d148ed2

Observation 6b8b0bce-50ce-4393-be4c-9ae5f8df441e · outbound

This paper cites an unresolved cited work.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:36:59.294077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.850582Z digest=sha256:9b961dc6519881c1d8f20fed749f0778495ad24c240e1a8eceb23e01e3bad71f

Observation f3967756-c0a0-4bf6-bb83-c4028c29d2c7 · outbound

This paper cites DeepDriving: Learning Affordance for Direct Perception in Autonomous Driving.ICCV, 2015.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts DeepDriving: Learning Affordance for Direct Perception in Autonomous Driving.ICCV, 2015

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.281067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.855606Z digest=sha256:fbc7bdb4f74858d3393f83021b2be7f77c57e038f656a1f8f9e44383fd038243

Observation ae499608-f5c4-4355-9d51-595d9c438154 · outbound

This paper cites Finding Safety Neurons in Large Language Models.arXiv preprint arXiv:2406.14144, 2024.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Finding Safety Neurons in Large Language Models.arXiv preprint arXiv:2406.14144, 2024

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:57.860304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:57.860304Z digest=sha256:d96ba023c5747c536a9820b3fddf35aded5dedc3748d8d5956e0ca6e3a588662

Observation fad7322a-2904-413d-a846-f0a374e9ed92 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:57.865187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:57.865187Z digest=sha256:2adc126f23b858fba34e4ce07480f06bd224f9dc7fd1b23cb052afb0ef5086e4

Observation 932cb4ec-7843-46b8-b39b-05682e2d5e9e · outbound

This paper cites Natural language processing (almost) from scratch.JMLR, 2011.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Natural language processing (almost) from scratch.JMLR, 2011

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.267292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.872050Z digest=sha256:ca51940f620ffc11a72a4baf11ededdd7f6f7bcace7e1c7ef28a97c1b4f569c4

Observation cc1c97ba-0ad4-4583-ab56-b3d1a3c3fae6 · outbound

This paper cites DeepSeek LLM: Scaling Open-Source Language Models with Longtermism.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:57.876865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:57.876865Z digest=sha256:cf95684346df76a69644b16cb08c35f3c7bfc6136b604895cdf720b77bc3965c

Observation f7328520-9651-4cc2-8ba3-00c44b3ec280 · outbound

This paper cites an unresolved cited work.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:36:59.254081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.882154Z digest=sha256:a5dbc51eaa56c5a42cfdafd565161c30df6c462dbc93df5ee165302bb08c5cfc

Observation 57aaf956-792b-4db4-9ad4-2785be8e1762 · outbound

This paper cites BERT: Pre-training of Deep Bidi- rectional Transformers for Language Understanding.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts BERT: Pre-training of Deep Bidi- rectional Transformers for Language Understanding

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.240256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.886676Z digest=sha256:04793da0b164de7eb546bcd2a9c785500731cf2b2c7620198c0a293267e8bd6e

Observation 42cb55e0-6667-4d2d-8093-0bfbe573e7a7 · outbound

This paper cites an unresolved cited work.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:36:59.226412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.891373Z digest=sha256:45acb64cfa0bbfd08ab0355299206105abd47fbf7ffd7c9087284cf8cd73e5be

Observation c4863368-b26d-416d-baef-3730e1746f6a · outbound

This paper cites Gemma 2 27B Instruction Tuned.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Gemma 2 27B Instruction Tuned

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.213377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.896914Z digest=sha256:ac69960476df7bea8c349ccb2f420e89849ed6b5b4414a9fdc54d0f1540f7ef2

Observation 30fc7dca-2646-47ee-8ab8-a4d3c2da007a · outbound

This paper cites Gemma 2 2B Instruction Tuned.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Gemma 2 2B Instruction Tuned

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.199591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.901664Z digest=sha256:3b9e288eaaa6ccc9333d4114c6ce2cccb0a19ee1f95a604abfd8f4a8b9ec4681

Observation 200dba93-d003-4897-8520-44ece877a7c7 · outbound

This paper cites Gemma 2 9B Instruction Tuned.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Gemma 2 9B Instruction Tuned

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.185368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.906402Z digest=sha256:038083ee75cc16397cb952f48a43ae5701d7f01f1ee676ec0e3b1c3068b816d8

Observation 1464180f-c9d3-4464-b93f-f2275742024a · outbound

This paper cites Gemma 3 1B Instruction Tuned.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Gemma 3 1B Instruction Tuned

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.171365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.911405Z digest=sha256:dab36dfdd9e5489a273c68142387a5258c6e396d0cac1ea2fc63c2be583bd48c

Observation bdd14b95-6036-4ae1-a189-075d134d296d · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Gemma: Open Models Based on Gemini Research and Technology

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:57.915684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:57.915684Z digest=sha256:5ba90353e8f10da45f3fb310a42f42df62beb5ea0721d0f48395a23f1325fe19

Observation 431344a7-d44f-4f3a-bf82-d595cbc63baf · outbound

This paper cites Qwen 2.5 14B Instruct.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Qwen 2.5 14B Instruct

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.157128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.920689Z digest=sha256:77afb80d9eafb30a3690a5f7dbb0abcc3fb3e6940e7179f9348d4380bd4de1e2

Observation b7c66642-f910-40c4-91bd-ab7449e7134e · outbound

This paper cites Qwen 2.5 32B Instruct.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Qwen 2.5 32B Instruct

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.142520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.925225Z digest=sha256:2ebd9fecd1f6f01643282d7e2794ebca6ca7011c239eb04652a6fcf2f3bb1fa4

Observation 4c59e1ed-02f6-457d-aa21-46947755431a · outbound

This paper cites Qwen 2.5 3B Instruct.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Qwen 2.5 3B Instruct

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.127811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.930164Z digest=sha256:ed8b383c64503ee8e0a71654b56a36eec266c2c40cd00c27f85f7b9ba5fb1a97

Observation 9c172ed8-f26e-49c7-a4e6-da1b4df6fe5e · outbound

This paper cites Qwen 2.5 72B Instruct.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Qwen 2.5 72B Instruct

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.110308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.934919Z digest=sha256:2e4c5a942f4d259989ca6cc4c64282ef03b2039acf81515c1f39399ce198a542

Observation 1d7656d8-32cf-4e92-9408-ccdd39a76e4f · outbound

This paper cites Qwen 2.5 7B Instruct.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Qwen 2.5 7B Instruct

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.095820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.939480Z digest=sha256:d2abebb4041af0bff50fd870781717fc24c17051901121a1ddcff6f2101d8ec1

Observation bbb9fb27-09b9-430d-a1b3-7d721c7a8dca · outbound

This paper cites BadNets: Identifying Vulnerabilities in the Machine Learning Model Supply Chain.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts BadNets: Identifying Vulnerabilities in the Machine Learning Model Supply Chain

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:57.943727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:57.943727Z digest=sha256:99284435a0ae38bef4151ee95ed8441429cac88197ca39a72dee0078c0d81fb0

Observation a738dcef-d76c-4ae6-be27-07bea86ff41c · outbound

This paper cites Gradient-based Adversarial Attacks against Text Transformers.EMNLP, 2021.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Gradient-based Adversarial Attacks against Text Transformers.EMNLP, 2021

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.082522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.949358Z digest=sha256:7a14afb28d07656125171cf35c6e8dee9c0d5d8846eb917c03ed161d082c7c72

Observation 64e0cd0e-73e9-4e3e-8b36-c9d254f5faad · outbound

This paper cites Hugging Face.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Hugging Face

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.069582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.954300Z digest=sha256:5155451e41a58fe95c0a0c42c2ffb2c1502845a06b89fca3227275632e2fbadb

Observation 748dbe20-cb50-49e2-a38a-ab0b7eac7bc3 · outbound

This paper cites Mistral 7B.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Mistral 7B

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:57.959074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:57.959074Z digest=sha256:55887ea69a1672845ed3727e49a57e20c52235082ae0da8ffbcf080ed4930dfc

Observation 7a98b255-044c-4f91-9418-8c11c29573d3 · outbound

This paper cites Kaggle: Your Home for Data Science.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Kaggle: Your Home for Data Science

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.056619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.963888Z digest=sha256:23a9833feacfeec0a5f210e24bef7c7c36e652f1df06a9144dc9e474047a9c33

Observation 43d0d9ca-796d-4538-9d5a-e15e259e16e5 · outbound

This paper cites Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks.IEEE SPW, 2024.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks.IEEE SPW, 2024

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.043152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.968436Z digest=sha256:d14ab9a9d621ffae57d2c729f2a4ceabe52a55455ac117c02ec3fbde1b1902c5

Observation 235d7e53-e848-4e48-a9ea-61b50c1785c7 · outbound

This paper cites TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts.USENIX Security, 2025.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts.USENIX Security, 2025

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.029104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.973333Z digest=sha256:626bdb71ff11192da06c8283573f576393572b4befd3e7632f26a37cfb53001d

Observation af9b9e39-8dc0-4736-8177-cb71e19cf6d6 · outbound

This paper cites SentencePiece: A simple and language-independent subword tokenizer and detokenizer for Neural Text Processing.EMNLP, 2018.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts SentencePiece: A simple and language-independent subword tokenizer and detokenizer for Neural Text Processing.EMNLP, 2018

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:59.016149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.978178Z digest=sha256:27c253567ef0476c88ec833dba95ddb20b4518b44c37acd6bfc3c6921aca2527

Observation 3ca1d9e0-2426-4105-a641-d716a1143aef · outbound

This paper cites an unresolved cited work.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:36:59.002599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.982656Z digest=sha256:0a01caeb3e2e32bdf5e8358fba186587839ac7b572d164c34164b4b02b47f79a

Observation a4bf452f-dc18-4ae3-a9c0-0c89bd18d61a · outbound

This paper cites Backdoor Learning: A Survey.IEEE Transactions on Neural Networks and Learning Systems, 2022.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Backdoor Learning: A Survey.IEEE Transactions on Neural Networks and Learning Systems, 2022

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.989392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.987594Z digest=sha256:4c6d869e7bea6ceba03339f86436d238a8ee960c7adf518699654186a4eec01e

Observation 46d91521-766a-42a7-8fc2-d4c8953ad6a5 · outbound

This paper cites AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models.ICLR, 2024.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models.ICLR, 2024

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.975576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:57.992623Z digest=sha256:03862ee97d99887f0d7ea449455b7d5bd562c23ffaa8df681fb989634c0cca6b

Observation b44450f2-c569-40fa-9b84-495b8ccdc68b · outbound

This paper cites Prompt Injection attack against LLM-integrated Applications.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Prompt Injection attack against LLM-integrated Applications

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:57.997650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:57.997650Z digest=sha256:0994c9bfa67b4e18ba2555b7482b26164c40ead42902cc618f888957fbfd1d37

Observation 419a786a-5820-4356-8525-bf33c695aa41 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:58.003299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:58.003299Z digest=sha256:9c1291e09c5e3035ebbb8af9570d16e5df7878b9205d1d500358585541ed9ef6

Observation c522ee51-582c-4774-8227-eea8ba257110 · outbound

This paper cites Llama 2 13B Chat.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Llama 2 13B Chat

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.960942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.008876Z digest=sha256:89c8a69a8353dfdff00503841a1dbbe3c2051b6e1145807a0f03e709ac7f4e81

Observation 1b48a51d-a368-4e5b-9b0c-4e9c641c6535 · outbound

This paper cites Llama 2 70B Chat.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Llama 2 70B Chat

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.947422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.013382Z digest=sha256:c64d6a6a633d69ed3d26b48e69ffcceb77991044a3e1f78dabb5e67c3337bdcf

Observation 5c1a7478-a653-487b-9896-6be0a84ccd9c · outbound

This paper cites Llama 2 7B Chat.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Llama 2 7B Chat

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.933489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.018626Z digest=sha256:528307e5818d57874afd77ea0215ddc7e6549fd305db09e07f4b629d05518a62

Observation 81429a98-3c39-4cd1-bb79-804aabad1ffc · outbound

This paper cites Llama 3.1 8B Instruct.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Llama 3.1 8B Instruct

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.919872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.023583Z digest=sha256:7bdeccb786694d9d4fdaddc7c18ede13bf7c4eb8eee7f2cdcb01de11f0ffa367

Observation 50edf83d-8106-4130-9231-a9a0265b565a · outbound

This paper cites Llama 3.3 70B Instruct.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Llama 3.3 70B Instruct

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.905409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.028769Z digest=sha256:1f850e04ad2b90d30ba6abcc8cc0417d6b0d70ab07efcbb227f9214c5214df1f

Observation 1d7dd1f0-a501-46db-a3bf-621a22213e6a · outbound

This paper cites Llama guard 3 8b.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Llama guard 3 8b

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.891368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.033447Z digest=sha256:bbdea0f119c66618b8ac4caffca8fcba7b9510a5c373d6115a04d74ece9ca1b2

Observation 87a12934-7594-48a7-8688-d537ea6f7c18 · outbound

This paper cites Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.877252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.038169Z digest=sha256:345d98eb793d0a816af76bdd49b2aec3a67d1c2afc3b2fd169fdfffcafbd46ac

Observation 0b2320a8-d09f-4b73-8953-30b083ddc985 · outbound

This paper cites an unresolved cited work.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Unresolved cited work

Reference 47

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T05:36:58.862074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.042801Z digest=sha256:2e94921b48634919a5f337d460702b8172e483bfebecf94a38bb1bf8c280858a

Observation 0854fd31-d5f6-4687-81c9-55f03057c7a1 · outbound

This paper cites GPT-4 Technical Report.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts GPT-4 Technical Report

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:58.047377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:58.047377Z digest=sha256:16247499c737cf946063c9605d7f56c8ff3845815e1c16e6c35cb4eee73b8d52

Observation 09546edd-6355-4941-9fbd-550e64cccb87 · outbound

This paper cites an unresolved cited work.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:36:58.847174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.053038Z digest=sha256:0e237edf4cecfe561e5997412d5042c49beedbe46b772f67ba3d0b2ae34bda1f

Observation 1dc111cf-1c33-438d-9d62-2d807d3fca6a · outbound

This paper cites Automated Red Teaming with GOAT: the Generative Offensive Agent Tester.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Automated Red Teaming with GOAT: the Generative Offensive Agent Tester

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:58.057842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:58.057842Z digest=sha256:7fb18a4c57d5767d3394156ee66b16e94032343d9586e76a372d4bf2c43a572b

Observation 1fd41837-867c-4e8b-9394-03d60eb485a8 · outbound

This paper cites Gradient Descent.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Gradient Descent

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.830890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.062774Z digest=sha256:73e9e004bbb9af1cc010c875e66b5195d93960bc01a497127ac3863e6ef2df04

Observation ed92a7e7-6207-4378-b555-6d3d17790395 · outbound

This paper cites an unresolved cited work.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:36:58.814992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.067654Z digest=sha256:e77b43a0c8550724a8756afd697f7a5252e5261796acf4e5a3a00d174aa8c9d8

Observation 56c857bd-92be-4fc8-ae18-dda51f1f0073 · outbound

This paper cites The Llama 3 Herd of Models.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts The Llama 3 Herd of Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:58.072430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:58.072430Z digest=sha256:585c8af41b3cc7f5d16c90f66a62989b31b355b5064175ea4d00c120546bb11c

Observation fd0fe736-56e2-4922-9f49-3afeba09c9cc · outbound

This paper cites WinoGrande: An Adversarial Winograd Schema Challenge at Scale.AAAI, 2020.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts WinoGrande: An Adversarial Winograd Schema Challenge at Scale.AAAI, 2020

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.800197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.077206Z digest=sha256:4edb5ba54c276b841399b745710e05b747fc987934ff17064f4e6bd1e43f2004

Observation 6a8ebfe2-3fcc-4bb2-be3f-749b3a309df3 · outbound

This paper cites Neural Machine Translation of Rare Words with Sub- word Units.ACL, 2016.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Neural Machine Translation of Rare Words with Sub- word Units.ACL, 2016

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.785772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.081943Z digest=sha256:ec1b2152e5123c687e9b7ebffc215cccbcd5fe1914450c5df397c23dac89f606

Observation c88a7991-a998-4bd2-bcc7-971a4a59e63b · outbound

This paper cites "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:58.086467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:58.086467Z digest=sha256:bf1f23778ad672b2d28d87ed7f0916be1fb4b20b1a89cfaab9678b16cd001c30

Observation dd6b4dc9-8314-4cda-85af-50ffce4abdf8 · outbound

This paper cites an unresolved cited work.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:36:58.770882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.091409Z digest=sha256:4e0666858bf12c1d31493332fd4c97f507c37c0ec84c978636f8e2d6329e1921

Observation 67c4dc04-fa81-445b-a13e-1a5e73ed3713 · outbound

This paper cites Gemma, 2024.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Gemma, 2024

Reference 58

Resolution
parse uncertain
no resolver link, observed 2026-08-07T05:36:58.095917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:58.095917Z digest=sha256:b81007c901961bc43bbabba672efc523ddf63b23b07a54ef5f8dcfee0c87a2d4

Observation fbdaf693-0569-42be-84f7-2153f8c0e3c8 · outbound

This paper cites Gemma 3, 2025.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Gemma 3, 2025

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:58.100779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:58.100779Z digest=sha256:d5d4a0212a0d2b4b1a013cfe487e9bd57aeea36a4fa157fb2445a8b51a4ea138

Observation b1d25097-5d47-4f4f-b981-161f50c9f53d · outbound

This paper cites Pytorch, 2022.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Pytorch, 2022

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.737769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.105708Z digest=sha256:cca38306e20fc03730d6721a7a7b2b302eccf466061bf27d3fc4d4bf25681226

Observation c603f4e8-17d7-4a05-89c1-5945d2ef7818 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:58.110690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:58.110690Z digest=sha256:4d42c9232271b95ea716d46347121413c7e24954b115c4b01c9bbd142d4ba053

Observation 5c1526ad-f5be-484f-98e7-2573dd02add4 · outbound

This paper cites Centrum voor Wiskunde en Informatica Amsterdam, 1995.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Centrum voor Wiskunde en Informatica Amsterdam, 1995

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.724100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.115965Z digest=sha256:ab582972eac386ee5a5a6fc36abead7035c99a145947a38e23e33093a93c3753

Observation a18405a5-5cfc-4230-9522-61f54cb7fb90 · outbound

This paper cites an unresolved cited work.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:36:58.708640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.120844Z digest=sha256:24793f049e389f8ab76e8a13475af54676f51149ef753664ef7e369cf13841ac

Observation 36195b31-1a83-48f9-9e28-ac244d5ee814 · outbound

This paper cites A Simple and Effective Pruning Approach for Large Language Models.ICLR, 2024.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts A Simple and Effective Pruning Approach for Large Language Models.ICLR, 2024

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.694479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.126111Z digest=sha256:24e3e46cc68daad80a96c299bd147c5677f7204e4708ac97bd145b34cbc37dab

Observation 5dd3082f-e0aa-48dd-a03a-1d576f8b1046 · outbound

This paper cites an unresolved cited work.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:36:58.680727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.131255Z digest=sha256:36405e76e64944e6dab418827c47b87eb5444e9f5405ecf7b4703d9821d6b3bb

Observation abd6921c-f91f-4c0e-b1af-47bf8baa381a · outbound

This paper cites Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:58.135736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:58.135736Z digest=sha256:347c81b6b5129589ce52471c981c1fc03b5b8bc49f68e1e37ece6b0fb5c86ffc

Observation fc70e041-e9e1-4943-9760-4bb594bbb20a · outbound

This paper cites an unresolved cited work.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:36:58.665664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.140272Z digest=sha256:5122e29eadf3852466eda466f7cba4605b254a42579bfd806138148ed1bbe761

Observation 7b986524-8426-48e8-97f4-b410817e86b7 · outbound

This paper cites Qwen2 Technical Report.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Qwen2 Technical Report

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:58.144812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:58.144812Z digest=sha256:8b546c8586a159edd2b628d6dafe699e6aacb166a8004e2c97f74b6909bfc031

Observation b633371f-6db0-4379-ac16-8d75271362fb · outbound

This paper cites an unresolved cited work.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:36:58.650035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.149800Z digest=sha256:b0bd114a9d152bb36a99ac9620170bc8ab3fb56a35e37b3cebd8aa88fe13a9e2

Observation e7482177-de9c-4389-931c-da572ce734cc · outbound

This paper cites NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning.AAAI, 2025.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning.AAAI, 2025

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.635315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.154405Z digest=sha256:46f2d15045b066b3ec3ce5788212560b2185c0a0b44afedc55458cc106bc3342

Observation 1286cca4-608d-45a7-baef-af7346ed68bf · outbound

This paper cites GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T05:36:58.160681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:58.160681Z digest=sha256:3bc4a5adaf7455e3fdab8d427b2a3416f5cc608cfe3552cc986932f80ceca9cd

Observation 3cb4a1b2-dfd7-4561-9f38-162b82ca55fc · outbound

This paper cites HellaSwag: Can a Machine Really Finish Your Sentence?ACL, 2019.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts HellaSwag: Can a Machine Really Finish Your Sentence?ACL, 2019

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.621030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.165507Z digest=sha256:67e3a6293a11b58e8984690e8612af3692f25177a373de5c3ee8a101c78db86a

Observation 09e75c66-8e2c-4d85-b9a2-3bdd3d676cf6 · outbound

This paper cites How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs.ACL ARR, 2024.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs.ACL ARR, 2024

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.606730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.170255Z digest=sha256:2387a09bc5d33bcb7a0ed064a72f79f8bc64fa3d1807d1192b5b08166d1d1e94

Observation daac7867-7860-4232-906f-d8f9a0e09b22 · outbound

This paper cites Understanding and Enhancing Safety Mechanisms of LLMs via Safety- Specific Neuron.ICLR, 2025.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Understanding and Enhancing Safety Mechanisms of LLMs via Safety- Specific Neuron.ICLR, 2025

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:36:58.590629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:36:58.174845Z digest=sha256:e66c26dc1442739485fd33292656a8757e073b2ba17965e0498e6ffb73e04161

Observation cf19b872-e853-43ed-807b-73791d8b8c34 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 75

Resolution
malformed identifier
no resolver link, observed 2026-08-07T05:36:58.179549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:36:58.179549Z digest=sha256:cffe4d1828276d962a05fbf35da7f3d85462afa9cb10555b5d9d585f0981667e

Pith citing papers

Observation 05abea8d-e50b-41c8-97f6-31c40d0436e9 · inbound

GoodVibe: Security-by-Vibe for LLM-Based Code Generation cites this paper.

GoodVibe: Security-by-Vibe for LLM-Based Code Generation TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T01:03:28.012736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:03:28.012736Z digest=sha256:2ab54675d1ac03d00c0b0c50530aaedb5c69d1d6d8a82505f1bca54e6a8d4e85

Observation e5ecc3ab-046e-43dd-a0ba-16e046acb8c0 · inbound

The Art of the Jailbreak: Formulating Jailbreak Attacks for LLM Security Beyond Binary Scoring cites this paper.

The Art of the Jailbreak: Formulating Jailbreak Attacks for LLM Security Beyond Binary Scoring TwinBreak: Jailbreaking LLM Security Alignments based on Twin Prompts

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:46:18.596218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T02:42:08.565972Z digest=sha256:f73ae188a2e2f0da5057c2a78af3f48ea099a0e6f64ad0fbbaff904c9a9a21f6