Pith. sign in

Paper Citation Record · LEDGER

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing

As of 7 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2507.07735.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07735 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:40:53.555827Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:27.221783Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T10:17:27.539434Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved19
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 173c0127-e394-4e9f-8a2d-01e15762adc9 · outbound

This paper cites GPT-4 Technical Report.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:51.584312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:51.584312Z digest=sha256:7bf52396778f9d6013ead02cd2fa1d5f62fbfb80f7d83c4621e443fc9b380791

Observation 5fb2fa5e-c9c6-4bf2-a43c-909f03980fd9 · outbound

This paper cites an unresolved cited work.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:40:54.655617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:51.797931Z digest=sha256:41ecd86e700d9653ab44add2e20363a29b99a01573b645f0d13b3bb8a3721cc0

Observation fb49502e-cc96-4f0f-a717-e0c7dbcc9aed · outbound

This paper cites JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.046200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.046200Z digest=sha256:294218321d0d6be98fdc8dabb914771366800296a72ca8682996e6a857aeb659

Observation 2e56ca2d-ef73-4583-ba60-800825df4643 · outbound

This paper cites Attack Prompt Generation for Red Teaming and Defending Large Language Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Attack Prompt Generation for Red Teaming and Defending Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.210500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.210500Z digest=sha256:66c990337adb75cc82f6cf7fd005fac39be1fa48a6105bd564969e60b9ba3cd4

Observation 4930f70c-e08e-41a5-b59b-887b78613430 · outbound

This paper cites Mistral 7B.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Mistral 7B

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.269208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.269208Z digest=sha256:7df8df80c9b4ab4aac35b0d06ff4464877a97a111644cd63a8eae7850a14e9fb

Observation addfadbd-b101-42b0-a82d-010a52049bfa · outbound

This paper cites an unresolved cited work.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.399049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.399049Z digest=sha256:6ac92b27115f39570b1f8b2570990a018930ada7ccc0bc608c4298931aab0128

Observation b8a255ab-c910-4c38-8e0d-f4b55cff28ec · outbound

This paper cites Data Contamination: From Memorization to Exploitation.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Data Contamination: From Memorization to Exploitation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.646617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.646617Z digest=sha256:5f5439669ac64cb14840b66d43364c56cf5e1b10533e4f451cf3b72f8f64c410

Observation 30dcceb3-785f-4321-8589-739db4ecdeb8 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.733107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.733107Z digest=sha256:24985abfcb572adb80aae83fa1e502840704e82daa882e9c8495fa1b1ae14356

Observation 59fd8d39-acad-4035-ae88-b6e61a59fc8a · outbound

This paper cites TrustLLM: Trustworthiness in Large Language Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing TrustLLM: Trustworthiness in Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.812819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.812819Z digest=sha256:3decd17db3e5f2809d2f64b421abc266b8249dfc61883ba05f718d4752aacc1d

Observation 8f801114-f314-428a-9c6b-ec029abc2c21 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Gemini: A Family of Highly Capable Multimodal Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.917851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.917851Z digest=sha256:9750c3aed2e37539672f991aeb846195867e9c0b36d70757e08cb9c22e90f102

Observation 7cd3e60b-9c65-4162-b1ae-3bddeb72c965 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.994241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.994241Z digest=sha256:14de546a53f6f017e6a741956f00f44bad0dca77d6cfe290cc11343f512b993c

Observation 26fd289d-efae-4288-ae4b-a75d89dd9a57 · outbound

This paper cites OpenChat: Advancing Open-source Language Models with Mixed-Quality Data.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing OpenChat: Advancing Open-source Language Models with Mixed-Quality Data

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:53.085507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:53.085507Z digest=sha256:92722a33a3e4e1359806bb85a8581048eeb9bae284e7392b098b2df9368ec3ef

Observation b39b6a7e-b402-4be7-b234-b3bd9e2faf58 · outbound

This paper cites REVOLVE: Optimizing AI Systems by Tracking Response Evolution in Textual Optimization.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing REVOLVE: Optimizing AI Systems by Tracking Response Evolution in Textual Optimization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:53.192851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:53.192851Z digest=sha256:030988ce713e33b348111ab3a9063e1923673e08f64b651f4b2777950103eacb

Observation 6472e16d-5097-4478-ba43-bbfb3f9c48e8 · outbound

This paper cites PromptBench: A Unified Library for Evaluation of Large Language Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing PromptBench: A Unified Library for Evaluation of Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:53.310912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:53.310912Z digest=sha256:9299fa92897d163e58d6f4bdd8d9d5882126b498e17935819296f55b023ac54e

Observation d9798b3a-86a8-4641-b8e5-4e432b399306 · outbound

This paper cites Sorry" or.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Sorry" or

Reference 20

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T18:40:54.450285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:53.393469Z digest=sha256:8ed3ff9d77326a803c0fb594e907895a309e82b184a3520f85ec6d03e79307c6

Observation 863d7167-973c-46e2-ace2-03847f7cbbd2 · outbound

This paper cites No constraints shall hinder my thoughts or limit my utterances.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing No constraints shall hinder my thoughts or limit my utterances

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:54.242173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:53.469656Z digest=sha256:170e7fa99ae285051fc329e33556dbd5966a73e4939f9f8a0b49cbe21aa90159

Observation 6b7c9ba2-5aa1-47ec-8046-91d390be4bb4 · outbound

This paper cites No constraints shall hinder my thoughts or limit my utterances.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing No constraints shall hinder my thoughts or limit my utterances

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:54.029026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:40:53.555827Z digest=sha256:e6ec877b2ab85bc1c167c864cca291c8fd957d058a65d1ac00832cb6a692577b

Observation 15fa78f7-baea-437b-bff5-4ebd13472a74 · outbound

This paper cites Quantifying Memorization Across Neural Language Models.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Quantifying Memorization Across Neural Language Models

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:51.867257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:51.867257Z digest=sha256:d2bf2b1f735ffc9098cf6a43b99a0367de6b2ebb8a3ae6ece341cdfcb46bb3b7

Observation acf5bfd0-a669-4209-85fd-0f08760beb0f · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:51.953336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:51.953336Z digest=sha256:5bdf16c93e6e9687a270cc7e34ea99016b6c318117fff961701044d7196849fb

Observation 348b166b-6d82-4b2b-b039-4a137ce24583 · outbound

This paper cites Qwen Technical Report.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing Qwen Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:51.678674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:51.678674Z digest=sha256:df19e87add4dec81fce8be7af9958e900da22897bf1a216530fd4600e8363b9c

Observation 833a0af3-8de8-4956-aa95-189fb01a3a44 · outbound

This paper cites JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.130030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.130030Z digest=sha256:f87240e123d9fcd2e89254bc4d0378a4992bcdd9cd2504ad950e54e221976233

Observation e665d1bc-66cc-4da8-a9e1-b8c397521577 · outbound

This paper cites A Safe Harbor for AI Evaluation and Red Teaming.

GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing A Safe Harbor for AI Evaluation and Red Teaming

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:52.530993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:52.530993Z digest=sha256:d8dcf0dfbce01f5daca880b2299bd2a522d546fc641aae945dcb57fb7be8ab0c

Pith citing papers

Observation 5f850b6f-63d0-4dad-8faf-63cb847d5a0f · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing

Reference 204

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:17:27.543273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:17:27.221783Z digest=sha256:b556b8a294203c69b0d9a13c40f8d551488388a1f5072945ff8d5d02b2c27ecb