Pith. sign in

Paper Citation Record · LEDGER

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks

As of 22 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 0 inbound Pith citation observations for arXiv:2608.00134.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.00134 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T01:16:15.489997Z

measured 66 of 66 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

66 of 66 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved66
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 798a21b2-7217-4fa7-b796-04582796ca4b · outbound

This paper cites GPT-4 Technical Report.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:09.264836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:09.264836Z digest=sha256:b6ec877f3a6433a742e4cd1d4b40bd669cde337c321cc016aaf5d940a006ed23

Observation 7e9400a6-d005-41c6-add6-9971e971b577 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Gemini: A Family of Highly Capable Multimodal Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:09.358040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:09.358040Z digest=sha256:efdd0d0b22a3bffdec73f16019cf2913dce31e4ee2a10faf41bdf49dbe57dc4a

Observation 197cda93-1f6e-451b-ba6b-589f4480435f · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks LLaMA: Open and Efficient Foundation Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:09.397390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:09.397390Z digest=sha256:6694be14970e72ad04bbf07b5e33b1a9ada7f347a3946d10b246c63dbd64719a

Observation 1afd53bc-d838-4a72-8bd4-2cb3df258454 · outbound

This paper cites A Survey of Large Language Models in Medicine: Progress, Application, and Challenge.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks A Survey of Large Language Models in Medicine: Progress, Application, and Challenge

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:09.467147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:09.467147Z digest=sha256:321a5b1da3a38945616e0ebaf1798399b967bd22097557cf8e523dbd1d00c3e2

Observation 155f7a77-09a2-4a6b-b4de-b421612b25a3 · outbound

This paper cites Exploring Recommendation Capabilities of GPT-4V(ision): A Preliminary Case Study.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Exploring Recommendation Capabilities of GPT-4V(ision): A Preliminary Case Study

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:09.515819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:09.515819Z digest=sha256:010945b2609406f9ebc634217b3116ceca7f21c6634d65b1d1867f44e9dc5dcb

Observation d1e15201-e534-458f-aad3-c61eaa37143d · outbound

This paper cites {LLM-Fuzzer}: Scaling assessment of large language model jailbreaks,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks {LLM-Fuzzer}: Scaling assessment of large language model jailbreaks,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:09.563310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:09.563310Z digest=sha256:f8f2223aec8a7496fb21ffbb5bab4b458d0ed71e1b703c0f4757028ceb7b3129

Observation bb0cadd4-0fc1-4fa6-ba87-396222896a5e · outbound

This paper cites Exploiting the Index Gradients for Optimization-Based Jailbreaking on Large Language Models.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Exploiting the Index Gradients for Optimization-Based Jailbreaking on Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:09.626615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:09.626615Z digest=sha256:d09477de5a412eee95125fdcfdbced30cfb20890578c42925765e93494912664

Observation ab6b3cef-02ed-44c9-aa4b-6c404eb1e212 · outbound

This paper cites Making them ask and answer: Jailbreaking large language models in few queries via disguise and reconstruction,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Making them ask and answer: Jailbreaking large language models in few queries via disguise and reconstruction,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:09.673549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:09.673549Z digest=sha256:10944e6e3bddd8952face5421099f84f0a57ad7f0e8bb33add3d3db41127dc1a

Observation 986c3a0b-3af2-4f0f-9beb-c9e6bdff0a39 · outbound

This paper cites Helping big language models protect themselves: An enhanced filtering and summarization system,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Helping big language models protect themselves: An enhanced filtering and summarization system,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:09.743694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:09.743694Z digest=sha256:b395403f79841ddc6431bac3cde9b98fd482ab5865c18a33426e2fd7a11d3109

Observation 404c6bc3-e694-4d71-8157-9785c1c560ff · outbound

This paper cites X-Teaming: Multi-Turn Jailbreaks and Defenses with Adaptive Multi-Agents.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks X-Teaming: Multi-Turn Jailbreaks and Defenses with Adaptive Multi-Agents

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:09.808925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:09.808925Z digest=sha256:a4efe7075ebee88f3a5f10069665ec2f83dcbd0f6fb7d563933fba4867f49068

Observation 42419e43-111b-46b5-987d-646f2b11d314 · outbound

This paper cites Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:09.952167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:09.952167Z digest=sha256:22d388e0809fc74812cc32981c0c7de3c92e292779888aeee46f17195cbf2506

Observation 37abdc89-91e0-4d96-b872-5b08253372c3 · outbound

This paper cites Datasentinel: A game-theoretic detection of prompt injection attacks,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Datasentinel: A game-theoretic detection of prompt injection attacks,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.059041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.059041Z digest=sha256:8f40f713f5bedb1c5fe0c350dca4ff4c3dd0531c8701a0c41cb63fb71c4ea963

Observation d9fddcb5-449b-45cc-b026-56190ae5a93b · outbound

This paper cites MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.099322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.099322Z digest=sha256:a4b6d2482ec123a2158fa071611fd004d66aa8f3bf988dbfe2970a9feb97691f

Observation 61e51d5b-99a8-4d04-bdca-02bb060bf374 · outbound

This paper cites Fight back against jailbreaking via prompt adversarial tuning,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Fight back against jailbreaking via prompt adversarial tuning,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.155570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.155570Z digest=sha256:983418f1f4bf57e39a1ae38a19d49f1fbd709055f6f4f684bbd5f4bbd53177d3

Observation 6b501988-70d3-46de-a021-4fa32eacd20a · outbound

This paper cites Safety-Tuned LLaMAs: Lessons From Improving the Safety of Large Language Models that Follow Instructions.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Safety-Tuned LLaMAs: Lessons From Improving the Safety of Large Language Models that Follow Instructions

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.204874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.204874Z digest=sha256:de8074029255d0a2f03b37d4ac0f0e60c48ced3ae775112782b704437f4f9991

Observation 7a4e43e2-0e5d-4de7-bdcb-23599c9c7c44 · outbound

This paper cites Distributional preference learning: Understanding and accounting for hidden context in rlhf,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Distributional preference learning: Understanding and accounting for hidden context in rlhf,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.255483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.255483Z digest=sha256:9059946384d4af78d2eb9ab976d7f6f3ed931b8eb2c7a70123fd5f0666575416

Observation c5232ba9-5422-459d-ba88-940486fce466 · outbound

This paper cites Attacking Large Language Models with Projected Gradient Descent.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Attacking Large Language Models with Projected Gradient Descent

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.330416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.330416Z digest=sha256:23bb0850216b511c901ad5e02cce1d83ec254b2c1a9d5bf0b55aa03d3655b5f1

Observation fa11cee1-adf0-4aaf-9489-059349dd9982 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.406423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.406423Z digest=sha256:2b5c908a7c401cd0475829cb78af75a960d302ef529495f50e2a8e29595b422d

Observation 0f7201d3-4551-4f8f-a646-d595f553792f · outbound

This paper cites AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.469528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.469528Z digest=sha256:9def74493e4edb5f1120ab63ae894c18d2b5e31f0f21c19a5641e21a7fc14623

Observation a36c6875-97c0-47a9-8bd8-25e0b0f97ecd · outbound

This paper cites AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks AdvPrompter: Fast Adaptive Adversarial Prompting for LLMs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.569808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.569808Z digest=sha256:328ada1d612535bd04b3b895b532a6578526ff0b3202b2ffd14b7240dba6e41b

Observation a37eead2-fe61-4568-bb8d-e8cebb9cacac · outbound

This paper cites How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.686916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.686916Z digest=sha256:36a3a8016bcf40729895b1e0a69954416377a627deff8f4400229f0ca73a9769

Observation 15e36a56-ffb3-45ed-9880-b7a8ba38d7b5 · outbound

This paper cites GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.758446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.758446Z digest=sha256:3a4c7b692be48daf70a411a22bb4cba1897f7af8f0f3a44cd1e80c8f40925da5

Observation b286444e-f098-4bb3-affc-415bf51b6227 · outbound

This paper cites Honeytrap: Deceiving large language model attackers to honeypot traps with resilient multi-agent defense,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Honeytrap: Deceiving large language model attackers to honeypot traps with resilient multi-agent defense,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.811489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.811489Z digest=sha256:e541922137359cc265bb25b6207662e265c87cbbbb5d895343910211f5a8a969

Observation 75924873-2eeb-426e-8578-a3e07ec5c661 · outbound

This paper cites Great, Now Write an Article About That: The Crescendo Multi-Turn LLM Jailbreak Attack.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Great, Now Write an Article About That: The Crescendo Multi-Turn LLM Jailbreak Attack

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.875255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.875255Z digest=sha256:7437e11aad148ab0851a86171d0dcfe36fc78aa211c9ecaf58e3b134aab32e4a

Observation f2c38328-d87e-4d06-bd39-e9b7c027bb6d · outbound

This paper cites Tree of attacks: Jailbreaking black-box llms automatically,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Tree of attacks: Jailbreaking black-box llms automatically,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:10.989189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:10.989189Z digest=sha256:2689e8c6dc3432b4186d495866b8a1e31c81d5bb7ab2852a51e35d5694d8ea41

Observation 698e5121-2960-4aa4-a4cb-2b58e6633457 · outbound

This paper cites PAPILLON: Efficient and Stealthy Fuzz Testing-Powered Jailbreaks for LLMs.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks PAPILLON: Efficient and Stealthy Fuzz Testing-Powered Jailbreaks for LLMs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:11.117001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:11.117001Z digest=sha256:a04b16d61b6d00ae06c65930d3228e40d088ed515f776e231c8a73196542e36b

Observation 6b57bb15-a2bf-48a0-aead-4011b5f951a5 · outbound

This paper cites ” do anything now.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks ” do anything now

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:11.197104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:11.197104Z digest=sha256:15736dfb3465a57b79498b894a119f5415fd28e1e829aa886885d65f6ec8c66d

Observation 1e0a7775-d09c-445a-b87c-59792a2412ed · outbound

This paper cites Bag of Tricks: Benchmarking of Jailbreak Attacks on LLMs.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Bag of Tricks: Benchmarking of Jailbreak Attacks on LLMs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:11.264145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:11.264145Z digest=sha256:b0914f6db0a75dbd5a0e87baec2e8cde2b41115d22ec29f86e0d55fc1ce6d99c

Observation ed4c2d9f-cf68-4431-816e-f8271dc0c47a · outbound

This paper cites A Survey on Trustworthy LLM Agents: Threats and Countermeasures.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks A Survey on Trustworthy LLM Agents: Threats and Countermeasures

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:11.444009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:11.444009Z digest=sha256:2603742031223f111f4173803f0ba5ddeae0cd38cc800898aef6a723c4533e56

Observation b701a1c7-75e3-470a-8697-8a2126d352c8 · outbound

This paper cites Masterkey: Automated jailbreaking of large language model chatbots,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Masterkey: Automated jailbreaking of large language model chatbots,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:11.540024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:11.540024Z digest=sha256:e9ed4bc558538ba83d6010035f18543156d357ad8d66fdba7c176fb02b7d22f2

Observation 9043d682-2db1-45fc-bc51-2d569acad0a4 · outbound

This paper cites Jailbreakzoo: Survey, landscapes, and horizons in jailbreaking large lan- guage and vision-language models,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Jailbreakzoo: Survey, landscapes, and horizons in jailbreaking large lan- guage and vision-language models,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:11.664743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:11.664743Z digest=sha256:f525c9ec477ddde4220d2e12d78c02c4968ae308f2638544c8ee7a4addd62bfc

Observation 1292fb00-17c5-4c13-b6d4-8fd53a9db6bc · outbound

This paper cites Gpt- 4 is too smart to be safe: Stealthy chat with llms via cipher,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Gpt- 4 is too smart to be safe: Stealthy chat with llms via cipher,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:11.766046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:11.766046Z digest=sha256:ddab5c1271ca1f63a20001bb38cd450a12470e1703dba914877ba4c07527b733

Observation e08f1a45-122a-409c-b3b1-99c7d60ad914 · outbound

This paper cites Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:11.894746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:11.894746Z digest=sha256:b532f427c9ee3969c6d55e049ad6bdfda567e0c67806eeb15f8e7c1be362f72c

Observation 43d83782-7851-4cf1-97a9-6032f993231a · outbound

This paper cites Boosting Jailbreak Transferability for Large Language Models.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Boosting Jailbreak Transferability for Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:12.004751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:12.004751Z digest=sha256:9266d3ff62591a643b34a7d042a72923ca9d805685d7b1747201694e881d0280

Observation be428ec5-f14b-4543-a2cc-a2b6aab04e02 · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:12.144831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:12.144831Z digest=sha256:f1315518404e107131ddd9d91f4d45e35034106797b57a51d4486218f3b33098

Observation 9481f41b-60c7-4599-9b37-10358127434f · outbound

This paper cites DeepInception: Hypnotize Large Language Model to Be Jailbreaker.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks DeepInception: Hypnotize Large Language Model to Be Jailbreaker

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:12.245800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:12.245800Z digest=sha256:a51c6e3f425fda638ae4f25041da1d956030b8bd0cc8f75608eca2d24cae762e

Observation bec4a2cb-5209-4cbe-9ebe-e9c849fb47d7 · outbound

This paper cites Tempest: Autonomous Multi-Turn Jailbreaking of Large Language Models with Tree Search.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Tempest: Autonomous Multi-Turn Jailbreaking of Large Language Models with Tree Search

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:12.344757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:12.344757Z digest=sha256:440e7128e5c9ce3d1160f384aa5c73ce5f8c4f57d65242db8b0349fec9091d3d

Observation 5b40f7fa-4ed7-4e8a-ad95-f43854b50ad1 · outbound

This paper cites Defending LLMs against Jailbreaking Attacks via Backtranslation.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Defending LLMs against Jailbreaking Attacks via Backtranslation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:12.515641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:12.515641Z digest=sha256:3811afa1e918304c1df098c333b2533ccb93ec95999b28816a69fb519dea5d1d

Observation d520fa31-9bdf-4ace-9186-f1ddc6454d72 · outbound

This paper cites Jailbroken: How does llm safety training fail?.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Jailbroken: How does llm safety training fail?

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:12.672284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:12.672284Z digest=sha256:e660cb209357ba824d89158e485814a0daa51c82e71336ec0009aa3e41af5974

Observation 5aa005cc-1aac-45f5-b27f-407bd964ef24 · outbound

This paper cites Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:12.719886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:12.719886Z digest=sha256:c33d247cf6c84d6f49528e546cc79e375fc33fff73bb7ce59993af36e71ef0f8

Observation 7a38e1ad-70e8-40d5-be3d-d16c2fb0d05d · outbound

This paper cites Attack prompt generation for red teaming and defending large language mod- els,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Attack prompt generation for red teaming and defending large language mod- els,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:12.788834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:12.788834Z digest=sha256:f187cb33f339865b52a6e58828768b13188633dc1ac7503500271aec6a28f1d6

Observation 48601139-73e1-473a-ad66-da94405e14d2 · outbound

This paper cites From Theft to Bomb-Making: The Ripple Effect of Unlearning in Defending Against Jailbreak Attacks.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks From Theft to Bomb-Making: The Ripple Effect of Unlearning in Defending Against Jailbreak Attacks

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:12.873272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:12.873272Z digest=sha256:1e6841acb0a15c68568d9478968b96507bb66bf8920d4f2f8548b0ffcdcc856d

Observation fc595a05-2b83-4101-a1c1-46c4445fa5de · outbound

This paper cites Baseline Defenses for Adversarial Attacks Against Aligned Language Models.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Baseline Defenses for Adversarial Attacks Against Aligned Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:12.949365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:12.949365Z digest=sha256:21a005bded0f95fc3b9dc44e83edcbe0a500b60833170ace1424b7efb7026a54

Observation 771bd9f5-e888-44a8-a10d-77ae5df37a46 · outbound

This paper cites Certifying LLM Safety against Adversarial Prompting.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Certifying LLM Safety against Adversarial Prompting

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:13.014747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:13.014747Z digest=sha256:58c23c7b16fbc0448351246b2ad71390877fd71e9f01275db4c6b872fbfd7921

Observation 81cb2df2-f37f-457d-a8d0-16e043624665 · outbound

This paper cites SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:13.134915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:13.134915Z digest=sha256:744f7a8b6346777d8daf44ad0b5cb77cd4cc77d75bf39c2cc703799004b86a5f

Observation ba8886ca-b654-4e61-8022-fc96553d7efc · outbound

This paper cites Defending Against Alignment-Breaking Attacks via Robustly Aligned LLM.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Defending Against Alignment-Breaking Attacks via Robustly Aligned LLM

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:13.252157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:13.252157Z digest=sha256:7c76e154e724d2d41e02344f337e01018f02732da84338213fc37a07f0253e51

Observation 9744ff32-f663-4470-a466-7f4a778c6c67 · outbound

This paper cites Defending chatgpt against jailbreak attack via self-reminders,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Defending chatgpt against jailbreak attack via self-reminders,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:13.364750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:13.364750Z digest=sha256:e1c7d28ceded888f9fceaa5192a8c115eae9a922a6cbc5d7c352d40db5b35841

Observation b0ef8a60-ca68-4930-a283-77ed33a62d93 · outbound

This paper cites Steering dialogue dynamics for robustness against multi-turn jailbreaking attacks,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Steering dialogue dynamics for robustness against multi-turn jailbreaking attacks,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:13.535734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:13.535734Z digest=sha256:942bb328e5bd8573f3f54ec0df58b223d456bd84de2afacea6989c8bf3e72c60

Observation 9921ae4a-ca99-4243-9263-0365000a2123 · outbound

This paper cites X-boundary: Establishing exact safety boundary to shield llms from multi-turn jailbreaks without compromising usability,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks X-boundary: Establishing exact safety boundary to shield llms from multi-turn jailbreaks without compromising usability,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:13.651328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:13.651328Z digest=sha256:48a9813f8f5770f682ee05a189eb039ce4d10fdc3bcef3860c6efd08a861f181

Observation b15ac3db-aece-42e1-aa4a-759eb576099d · outbound

This paper cites RED QUEEN: Safeguarding Large Language Models against Concealed Multi-Turn Jailbreaking.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks RED QUEEN: Safeguarding Large Language Models against Concealed Multi-Turn Jailbreaking

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:13.794829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:13.794829Z digest=sha256:7a9e71a193bc14cbbf4bd02fb10dd5b85fed8d0c2a2b946fe1292c24a5d79291

Observation 1e43c751-4a75-44ae-8c4a-12950bd73985 · outbound

This paper cites Generative agents: Interactive simulacra of human behavior,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Generative agents: Interactive simulacra of human behavior,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:13.872894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:13.872894Z digest=sha256:ae4f897ab45873409cf067bb60ef724ed948ab15e7569db6336144e19b3a5f16

Observation 6bb1203a-fa28-45f2-88ae-bd7312b242a7 · outbound

This paper cites Training Socially Aligned Language Models on Simulated Social Interactions.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Training Socially Aligned Language Models on Simulated Social Interactions

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:13.946073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:13.946073Z digest=sha256:62c999a9fcb520fa9ac6437d0456544a9def2eeeef83401ea73e57a9995b2ab8

Observation ac613f5c-476e-4f49-936a-dd1cf9cfb4f0 · outbound

This paper cites Camel: Communicative agents for.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Camel: Communicative agents for

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:14.022585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:14.022585Z digest=sha256:08c61be2ea9884e07cb1e140b251a1537a290b6b270c28b938d0d090ce77efb0

Observation 8a872541-14d7-4d58-8867-cd96d55ade56 · outbound

This paper cites AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:14.080163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:14.080163Z digest=sha256:242113074d3a999af0db6f0b8f55c24aa0c238a7f27ff722d451feaf10e32a95

Observation d3c7f28f-25fa-4bd5-9330-81387c2e42df · outbound

This paper cites Metagpt: Meta programming for a multi-agent collaborative framework,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Metagpt: Meta programming for a multi-agent collaborative framework,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:14.172690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:14.172690Z digest=sha256:9688b89b3ec1828009121aa161229440887f6234dc29bf5fa019a132b8e5a508

Observation 66ee22d6-3d85-4913-9552-7bfac0e5507c · outbound

This paper cites ChatDev: Communicative Agents for Software Development.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks ChatDev: Communicative Agents for Software Development

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:14.275406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:14.275406Z digest=sha256:98941776db2c9e6c6925777a4ba6b0f31df8382aaf4e9f73510846e78bcf3497

Observation 8892c0c9-d9e9-4521-b8d7-f7b5f9749de4 · outbound

This paper cites Improving Factuality and Reasoning in Language Models through Multiagent Debate.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Improving Factuality and Reasoning in Language Models through Multiagent Debate

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:14.424749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:14.424749Z digest=sha256:6834bc62b77d78ed7dc70b48712d5c4e66a5e27f0bb9e5135c96e33fc395b9a4

Observation c713df20-a4d7-48cd-9945-5ffdffa5331d · outbound

This paper cites Encouraging Divergent Thinking in Large Language Models through Multi-Agent Debate.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Encouraging Divergent Thinking in Large Language Models through Multi-Agent Debate

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:14.553943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:14.553943Z digest=sha256:d8532fa8e53cdca3cac1fdfba30486c0872dde2edfc157e8025d1d6a5384e98b

Observation ccbcfc1e-1960-4ffa-af48-ee77962be44e · outbound

This paper cites Defending large language models against jailbreaking attacks through goal priori- tization,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Defending large language models against jailbreaking attacks through goal priori- tization,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:14.624835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:14.624835Z digest=sha256:9fdb9be7c9a30d2c7e221f50d8675d4c27849daf22e05334ad82297feed9b831

Observation 3de10f62-1386-4ca4-9a38-e8c914f54690 · outbound

This paper cites Robust prompt optimization for defend- ing language models against jailbreaking attacks,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Robust prompt optimization for defend- ing language models against jailbreaking attacks,

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:14.814826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:14.814826Z digest=sha256:46ec834f98213220ac2ad2c6d8392cbb409ae9efad5d117271d8e1957e1dffbb

Observation e49c6208-7c87-4509-8609-7c19ad7b1c60 · outbound

This paper cites SecurityLingua: Efficient Defense of LLM Jailbreak Attacks via Security-Aware Prompt Compression.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks SecurityLingua: Efficient Defense of LLM Jailbreak Attacks via Security-Aware Prompt Compression

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:14.920942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:14.920942Z digest=sha256:5163bf7caa753b4d059af07c395a6ac99dcbf923225c71fd4c23cfac244f2101

Observation ad2bd6db-372d-4528-bcb9-99a440857d4b · outbound

This paper cites Fine-tuning aligned language models compromises safety, even when users do not intend to!.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Fine-tuning aligned language models compromises safety, even when users do not intend to!

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:15.024609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:15.024609Z digest=sha256:308083d2bb855fb6125367b67836a7ec79c7b7564669ea1316e477eac2040b51

Observation cfe5471f-fd88-46ed-a81b-d8486bf54dc1 · outbound

This paper cites Jailbreaking black box large language models in twenty queries,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Jailbreaking black box large language models in twenty queries,

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:15.148273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:15.148273Z digest=sha256:80adefb91b5586d2468d9ac70c409dd2727305a80beb42ace4ed2f6fbfb1d4aa

Observation d2829aa8-a185-4f8d-bd83-8e7ce496b17e · outbound

This paper cites M2s: Multi-turn to single-turn jailbreak in red teaming for llms,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks M2s: Multi-turn to single-turn jailbreak in red teaming for llms,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:15.272601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:15.272601Z digest=sha256:c5e75dfbacc2da8083cd57417d8f37c0bc5c63da7eefdafa24e55af055de510f

Observation 0a22fc85-a5fe-456f-b43a-4d89dc2add11 · outbound

This paper cites MT-Bench-101: A Fine-Grained Benchmark for Evaluating Large Language Models in Multi-Turn Dialogues.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks MT-Bench-101: A Fine-Grained Benchmark for Evaluating Large Language Models in Multi-Turn Dialogues

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:15.396819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:15.396819Z digest=sha256:804f1fbccbc851bb9d4c1c067866f17be7460cdfc2ed8d757a850bbf0d668209

Observation 1702fe08-30d1-48ab-94b9-b7ac696391da · outbound

This paper cites Cosafe: Evaluating large language model safety in multi-turn dialogue coreference,.

Stateful Cooperative Agents Safeguarding LLMs Against Evolving Multi-Turn Attacks Cosafe: Evaluating large language model safety in multi-turn dialogue coreference,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-04T01:16:15.489997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:16:15.489997Z digest=sha256:1f2e0c9a0c605895ce1f853a967846c1cb66f741b45de9a2ed2e282176a517b7

Pith citing papers

No inbound Pith citation observations are available.