Pith. sign in

Paper Citation Record · LEDGER

JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2402.05668.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.05668 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:52:15.472625Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T19:36:08.555578Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a9b03fa3-e25e-4aab-91d9-3a40921e7934 · inbound

Refusal in Language Models Is Mediated by a Single Direction cites this paper.

Refusal in Language Models Is Mediated by a Single Direction JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 126

Resolution
verified exact
arxiv_id, observed 2026-05-13T10:47:56.068686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T10:47:55.934081Z digest=sha256:e4f65bacb427c06d31ca56e773d9651c5b8629640d492d1bf797f8e34d9602eb

Observation 73c472e5-a747-4521-b49c-e63f8ad69b10 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.591202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:7788e239ab0c2a89314d459d9ee77c85d5f761c72202f4c6123c4f3aa9e28199

Observation 4efbf67c-a5e7-4ffe-9e72-3869d0b34bee · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:26.398289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:ba4babc78b7b8e240379ef5d7127bae9eb99b7408493158a2ffbee6e63d4d03d

Observation 9220165c-0f83-4a54-8e6b-180be994e0c6 · inbound

Towards medical AI misalignment: a preliminary study cites this paper.

Towards medical AI misalignment: a preliminary study JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:15.472625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:52:15.472625Z digest=sha256:9e09ae4a43ad14cf6ecad952d1c223267b46e3ea31dbc387b3695d329f7b5b92

Observation 704992cc-917f-4293-987e-29d345098244 · inbound

Audio Jailbreak Attacks: Exposing Vulnerabilities in SpeechGPT in a White-Box Framework cites this paper.

Audio Jailbreak Attacks: Exposing Vulnerabilities in SpeechGPT in a White-Box Framework JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:27:05.591822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:27:05.591822Z digest=sha256:dc43368fa0bc5d01719625bc83b0fb36e625aa53b8c6332848a7b89f679b4c0b

Observation b2eaf344-2622-4f66-b217-4850c0eb5904 · inbound

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures cites this paper.

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:40:20.651700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:40:20.651700Z digest=sha256:e9b445a03e03fb065a7f07436ef6c3d14149bbe069bdf9d3322ce3cfca364869

Observation 266333ba-7c41-48f7-85d0-e27a04f77d93 · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:26.788943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:26.788943Z digest=sha256:f500d7012a82b56ad1eb8fa49b2ca4d958a17c4474ba524aeb545517107ad13a

Observation f4547e9d-2bed-49f3-a872-041039dcc0f2 · inbound

Investigating Vulnerabilities and Defenses Against Audio-Visual Attacks: A Comprehensive Survey Emphasizing Multimodal Models cites this paper.

Investigating Vulnerabilities and Defenses Against Audio-Visual Attacks: A Comprehensive Survey Emphasizing Multimodal Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:08:39.731690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:08:39.731690Z digest=sha256:b689fcfb500a3990149761c609a9299950b3d96738fec9766c36c4074ef20bb2

Observation 92da5a23-91f1-46d2-80c5-233ed4293158 · inbound

InfoFlood: Jailbreaking Large Language Models with Information Overload cites this paper.

InfoFlood: Jailbreaking Large Language Models with Information Overload JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T01:02:28.442955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:02:28.442955Z digest=sha256:6aae19a6f4a6d6a4ba024962cf8e8bee6122ae45cd6312965bbc7de89d937549

Observation 53031d22-f6a6-49d8-938f-7cc3cf6f8c47 · inbound

Q-resafe: Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models cites this paper.

Q-resafe: Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:19.513275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:00:19.513275Z digest=sha256:3892d80db233d61eef2271486b3bce7399c22cfcf836242df967bbd7f1e7125c

Observation 483123a3-0bfe-4cc5-9c74-2c4268e7b18d · inbound

Understanding How University Guidelines Address Privacy and Security Issues of Generative AI in Academic Settings cites this paper.

Understanding How University Guidelines Address Privacy and Security Issues of Generative AI in Academic Settings JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T22:49:58.020442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:49:58.020442Z digest=sha256:0e2c7cdba887c607aa0f318e3fccd4f7021bce934a9a26a0d03874163d77632f

Observation 01705d48-dbcf-4a46-bc10-2a5b4228527b · inbound

Linearly Decoding Refused Knowledge in Aligned Language Models cites this paper.

Linearly Decoding Refused Knowledge in Aligned Language Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:27:34.892626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:27:34.892626Z digest=sha256:a80cebfbe16a8a0bb1981fbdaa3d97b5881648e503b15abb522b433c12a1d0d5

Observation 155f482f-eb0c-4b09-930b-c5236f1f9e56 · inbound

CAVGAN: Unifying Jailbreak and Defense of LLMs via Generative Adversarial Attacks on their Internal Representations cites this paper.

CAVGAN: Unifying Jailbreak and Defense of LLMs via Generative Adversarial Attacks on their Internal Representations JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:18:42.941390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:18:42.941390Z digest=sha256:7d816580776aa8b40d42f9f734394167484d2945ddcf1237ac09f2f1d7220abf

Observation 833a0af3-8de8-4956-aa95-189fb01a3a44 · inbound

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing cites this paper.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.130030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.130030Z digest=sha256:f87240e123d9fcd2e89254bc4d0378a4992bcdd9cd2504ad950e54e221976233

Observation a1426feb-59eb-4b50-8b38-4e88560801cd · inbound

An Audit and Analysis of LLM-Assisted Health Misinformation Jailbreaks Against LLMs cites this paper.

An Audit and Analysis of LLM-Assisted Health Misinformation Jailbreaks Against LLMs JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T01:02:13.588479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T01:02:13.588479Z digest=sha256:3cd3906eeb0604152a0bc504b8cf744e4cfd7722e172dd00679dfc7d4704683a

Observation 8d1dba16-b206-4dc7-b45c-769d6d1a2b67 · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 135

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:45.234154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:45.234154Z digest=sha256:a93a052cc9d86dc18962c33a6f484da3fb512e54400d64327cb31a7160f64071

Observation 3c94b144-c84e-4a7f-92de-00f9a3c5adf4 · inbound

ORFuzz: Fuzzing the "Other Side" of LLM Safety -- Testing Over-Refusal cites this paper.

ORFuzz: Fuzzing the "Other Side" of LLM Safety -- Testing Over-Refusal JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T23:31:54.568455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T23:27:52.438709Z digest=sha256:2306bcbda39626625b8be75beac282b527c2ca23057223388a3761f23bba4031

Observation b71221d5-a147-49b2-8228-46fa8024f7d3 · inbound

SALMAN: Stability Analysis of Language Models Through the Maps Between Graph-based Manifolds cites this paper.

SALMAN: Stability Analysis of Language Models Through the Maps Between Graph-based Manifolds JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T17:14:37.440851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:14:37.440851Z digest=sha256:1cb6350be9f1dabc5f52e9ebb1e9d2c2696f557138d3f90842e60b76cdc82aad

Observation 4cd21b06-c7ea-41e7-b016-ad9b2b1c5695 · inbound

Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection cites this paper.

Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T14:57:12.791829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:57:12.791829Z digest=sha256:c69343bafe20f222c2a9f00e9da8bd1144e018564b8ff8b79e05d59570d2c3ca

Observation 914ffc20-a3ee-46f8-9239-f750fb86e931 · inbound

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring cites this paper.

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T14:51:03.645921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:51:03.645921Z digest=sha256:dc705982240e5ca485d7d3c68adc0f8d1370ac5e7e291e776d35b75d283087e5

Observation e95fd325-7c11-47a5-ae6b-5c7da99a67f0 · inbound

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems cites this paper.

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T17:46:15.475308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:46:15.475308Z digest=sha256:e88b1bdeffa43c0b191d85a687aa89dc2b8252b0a3b95973139b75411e1746fd

Observation 4efc1b01-efce-4b62-be4b-0c1e0775fafb · inbound

How Well Do AI Systems Solve AP Physics? A Comparative Evaluation of Large Language Models on Algebra-Based Free Response Questions cites this paper.

How Well Do AI Systems Solve AP Physics? A Comparative Evaluation of Large Language Models on Algebra-Based Free Response Questions JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-15T13:15:31.225992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:15:31.225992Z digest=sha256:dec00a1b5dfe3316380c2c5f2f1eced6861316b625673ef5ff3e50b835ffaa0b

Observation 83b253df-03b6-4e6f-9d82-15dc863af4a2 · inbound

The Art of (Mis)alignment: How Fine-Tuning Methods Effectively Misalign and Realign LLMs in Post-Training cites this paper.

The Art of (Mis)alignment: How Fine-Tuning Methods Effectively Misalign and Realign LLMs in Post-Training JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:45:50.609141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T18:18:56.476698Z digest=sha256:1aa7c3a0c5cf103aa83a6172bf3375409a4e473e6fbab1de3c886ebe5811f6b3

Observation bb8ba95f-ef39-43f5-b0d1-985a43a95377 · inbound

SoK: Robustness in Large Language Models against Jailbreak Attacks cites this paper.

SoK: Robustness in Large Language Models against Jailbreak Attacks JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:01:08.866143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T16:42:41.137808Z digest=sha256:669e97d6a396c421319724822da57bea27e66ed3cc5822a786489d58e11a26da

Observation 356aab86-984e-425c-80a4-1165029df8ba · inbound

Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Models cites this paper.

Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:36:08.557072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T22:23:00.894059Z digest=sha256:ceb7c2b8081f5df0b01321fca513c3aa1dd99d50e7fee1302a3a8176b5e1296b

Observation 6453eee7-1ef5-458a-9d47-4ea67e510ce6 · inbound

SCARCE: Scalable Cascade Analysis for Rare-event Characterisation via Embeddings cites this paper.

SCARCE: Scalable Cascade Analysis for Rare-event Characterisation via Embeddings JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:04:20.789297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T07:03:36.414202Z digest=sha256:9b879b704004f9c6dcfc07a63a42edba1cb159b68b9ebe81fd4528b3920733e5

Observation 30413506-c871-4d98-ae6d-1776f7b318c1 · inbound

AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models cites this paper.

AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T08:30:02.436170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:30:02.436170Z digest=sha256:50ea472bac0a1d0f27b9b3323ca1dce93309e367c542f0f559c110644328224c