Pith. sign in

Paper Citation Record · LEDGER

The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2401.00287.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.00287 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:28:00.782958Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T13:28:22.268414Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 19caa072-27a4-4f51-85ed-b3d93f458b3a · inbound

No Free Lunch for Defending Against Prefilling Attack by In-Context Learning cites this paper.

No Free Lunch for Defending Against Prefilling Attack by In-Context Learning The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T15:49:38.125760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:49:38.125760Z digest=sha256:7bb28bc48686fb23a6731bb6bb68c0ec09ca5eea62fb2158651f904a6fc97fdd

Observation c05bef07-2c6a-456a-982d-140b59fd6193 · inbound

Token Highlighter: Inspecting and Mitigating Jailbreak Prompts for Large Language Models cites this paper.

Token Highlighter: Inspecting and Mitigating Jailbreak Prompts for Large Language Models The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T05:02:26.704127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:02:26.704127Z digest=sha256:1169a09b80cfd24ebf5560f1fe24e2d187fda2ea044d937ff91b36be95191dea

Observation 913e0770-a6b9-4126-8d9c-28046a78c6c2 · inbound

Aegis2.0: A Diverse AI Safety Dataset and Risks Taxonomy for Alignment of LLM Guardrails cites this paper.

Aegis2.0: A Diverse AI Safety Dataset and Risks Taxonomy for Alignment of LLM Guardrails The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T20:15:54.842752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:15:54.842752Z digest=sha256:bb52ac52563c5969ad3ce4f06ffee42bf164a5713baea31b84597bed5c2c4b8a

Observation 01e275ba-f362-444f-95ad-03271447018e · inbound

The TIP of the Iceberg: Revealing a Hidden Class of Task-in-Prompt Adversarial Attacks on LLMs cites this paper.

The TIP of the Iceberg: Revealing a Hidden Class of Task-in-Prompt Adversarial Attacks on LLMs The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T13:53:10.980458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T13:53:10.980458Z digest=sha256:f2215d2abc5341dd1484c64aea297f6b4657b9dbe45399e39925f129df158b2d

Observation 4adf32af-18be-409b-94e6-c2ae08f883d5 · inbound

`Do as I say not as I do': A Semi-Automated Approach for Jailbreak Prompt Attack against Multimodal LLMs cites this paper.

`Do as I say not as I do': A Semi-Automated Approach for Jailbreak Prompt Attack against Multimodal LLMs The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T17:58:57.515461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:58:57.515461Z digest=sha256:47d1c899654e5c0de2cebe19079cdd62fa450f6e8e72d381c00c45cdc2d88f10

Observation dfe3773f-abb7-4a73-9c47-d90c3ee23680 · inbound

GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation cites this paper.

GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T17:31:59.882927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:31:59.882927Z digest=sha256:707dcd82ae6fa27b5728ce1e011d63ae69a448e3a2c4953ecad56a5275a9891d

Observation 0a5407f4-98b3-45da-aef6-3b1e9f8c7124 · inbound

LLM Security: Vulnerabilities, Attacks, Defenses, and Countermeasures cites this paper.

LLM Security: Vulnerabilities, Attacks, Defenses, and Countermeasures The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness

Reference 135

Resolution
unresolved
no resolver link, observed 2026-08-16T04:28:00.782958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:28:00.782958Z digest=sha256:d73d8774510bd987daf2e98e6cf6ab5dec185b64a10a09bcf45b874a4fdf6736

Observation 20202f5c-7b09-449a-bf81-997ce875bb6f · inbound

System Prompt Extraction Attacks and Defenses in Large Language Models cites this paper.

System Prompt Extraction Attacks and Defenses in Large Language Models The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:28:22.318470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T13:28:18.665210Z digest=sha256:d28283a16e1f4c8492d3bf46bf739aa533a596c00b5b652d9156c96a98111a58

Observation a35ca690-ee7d-4ed1-9390-0c435e34c5e2 · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:26.690689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:26.690689Z digest=sha256:d003842aa0c6b3633e5fe171fb279afbe5c7ce8fabcafc50d695258ff5730dcb

Observation df5d9952-ca8b-4087-abda-50187b3a32d7 · inbound

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems cites this paper.

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems The Art of Defending: A Systematic Evaluation and Analysis of LLM Defense Strategies on Safety and Over-Defensiveness

Reference 295

Resolution
unresolved
no resolver link, observed 2026-08-04T17:46:19.737771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:46:19.737771Z digest=sha256:81f8cd2f7b5cca1e93b23f148b07cdcee0b3d669eff2335ac2552c758331b0a2