Pith. sign in

Paper Citation Record · LEDGER

GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2402.13494.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.13494 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:56:34.413670Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:29:42.339750Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bfc5b8be-fe6f-42bb-9849-bbbbf7bab971 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 101

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.432029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:ec7d8bef9eed4b9d5a1a761cd01019ca68fedcf766042f1b684049620f1eb119

Observation c628ec6d-4b6e-4b2d-821f-4fe4f4b1549c · inbound

Preventing Jailbreak Prompts as Malicious Tools for Cybercriminals: A Cyber Defense Perspective cites this paper.

Preventing Jailbreak Prompts as Malicious Tools for Cybercriminals: A Cyber Defense Perspective GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T12:56:34.413670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T12:56:34.413670Z digest=sha256:bff143c877f9f4853d5daebbb06eaf43101dd2238a8f6346c7473ad221a0b4e4

Observation b165e047-38ca-4c02-b6e1-6c0447bc211c · inbound

JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation cites this paper.

JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T12:25:30.681565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:25:30.681565Z digest=sha256:2cbf877e9967e82d8d5d81db0da8a6f66504ef815c2c2c56496049db80a895d5

Observation d81d0741-e5fd-4645-8705-a26bd19019b9 · inbound

The First Differentiable Transfer-Based Algorithm for Discrete MicroLED Repair cites this paper.

The First Differentiable Transfer-Based Algorithm for Discrete MicroLED Repair GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T22:21:02.842889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:21:02.842889Z digest=sha256:5f351e3faba2b4cfabece15d11c9e38db9f84108bcfd826596cf202642b90226

Observation f04e810c-0149-4d36-a99b-ea39639eca8f · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 202

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:58.003099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:58.003099Z digest=sha256:13e1a4dd27604591d2d603c24a0ad26a3d2208ef23ee2cb67824c906981e62e9

Observation fdb5ed4b-ee3e-45cb-ad40-d7363ff7f1aa · inbound

Distributionally Robust Token Optimization in RLHF cites this paper.

Distributionally Robust Token Optimization in RLHF GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:23:16.015683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-14T23:19:47.586351Z digest=sha256:d199f74ed1186b4aba15ba826626688c618af3a6eb1a84861bf0cc1786254d5f

Observation dffa1a2a-ab29-4773-ad4c-e83dab82cdf2 · inbound

SafeAgent: A Runtime Protection Architecture for Agentic Systems cites this paper.

SafeAgent: A Runtime Protection Architecture for Agentic Systems GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:11:20.543560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T06:06:49.717091Z digest=sha256:38aa906a988a5db5b9b2734eba70cfbff40bcd683cf38d1ff8ff8eca30c771ee

Observation cd7b3706-ae23-4913-9dce-efa232d47631 · inbound

Defending Jailbreak Attacks on Large Language Models via Manifold Trajectory Kinetics cites this paper.

Defending Jailbreak Attacks on Large Language Models via Manifold Trajectory Kinetics GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:37:14.900803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-27T21:55:48.561400Z digest=sha256:abc509359550a9f3f7840759877ea4764fc85f03c453ed247ea2f7bbaef14dd8

Observation 7c183d37-1553-43ec-8b63-7a205ef9d94b · inbound

Investigating The Security of Modern AI and Cloud Infrastructure cites this paper.

Investigating The Security of Modern AI and Cloud Infrastructure GradSafe: Detecting Jailbreak Prompts for LLMs via Safety-Critical Gradient Analysis

Reference 142

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:29:42.341232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T11:31:39.910784Z digest=sha256:7597240da84ce934fd794d6fd1598e314e103b7ed14fc15f3a79b93c26a7079e