Pith. sign in

Paper Citation Record · LEDGER

PAL: Proxy-Guided Black-Box Attack on Large Language Models

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2402.09674.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.09674 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:57:58.326747Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T11:57:03.488334Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation efeb32c0-20bc-4c05-ad0c-89ff25a44a81 · inbound

JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models cites this paper.

JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:08:05.491124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-15T06:08:05.386345Z digest=sha256:b66ecc5a0f0dd84d9a21efabb7153d2734df580b8a2cc13e1756ebd0ae600603

Observation 4ca240ad-03f0-4bdb-be64-fdbb4e2df369 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:20:44.829075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:151f81556e96f7dcb289ac535a9ad0dc332d68b14427daee2ff4d44dcc0ce483

Observation 04b34d31-aae2-4e1c-8588-f591b0453695 · inbound

Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment cites this paper.

Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T11:03:01.222327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:03:01.222327Z digest=sha256:80b85a5f25d92e347d99a7983244f18e0a3e5020c4495e13f41fbda5ab99da5e

Observation 5f759c0e-d78e-4acf-ae46-a22f5a8c85c4 · inbound

Does Safety Training of LLMs Generalize to Semantically Related Natural Prompts? cites this paper.

Does Safety Training of LLMs Generalize to Semantically Related Natural Prompts? PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T22:41:27.765662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T22:41:27.765662Z digest=sha256:3fbcacc9646c021ba25ec756e3c771433b2ab7f9b7f33abaa2eb2bb02ffd3c46

Observation ac98da78-1312-49dd-b009-17cba532858a · inbound

LIAR: Leveraging Inference Time Alignment (Best-of-N) to Jailbreak LLMs in Seconds cites this paper.

LIAR: Leveraging Inference Time Alignment (Best-of-N) to Jailbreak LLMs in Seconds PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T20:53:37.881247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:53:37.881247Z digest=sha256:a7e4d5e879284f3a517d8fc1bd5bbde9781f42b01eb62b82c377264be8570c48

Observation 1a297fb3-6279-4d64-883d-86c94f074646 · inbound

Diverse and Effective Red Teaming with Auto-generated Rewards and Multi-step Reinforcement Learning cites this paper.

Diverse and Effective Red Teaming with Auto-generated Rewards and Multi-step Reinforcement Learning PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T04:38:36.695356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:38:36.695356Z digest=sha256:9070f2c09bd59dfdf2fc51449c75514729de47b47507a61d98a04f7ca99d12de

Observation 48c1c031-ad95-4aaf-819a-6ec6b509640b · inbound

On the Validity of Traditional Vulnerability Scoring Systems for Adversarial Attacks against LLMs cites this paper.

On the Validity of Traditional Vulnerability Scoring Systems for Adversarial Attacks against LLMs PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 126

Resolution
unresolved
no resolver link, observed 2026-08-10T23:40:19.074950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T23:40:19.074950Z digest=sha256:c0499d976965813624bc9f91c4a8baf37ddd78a33c25122a418868c5f7b7a39d

Observation 70f2dad5-0b57-4929-a353-aebc6abe5965 · inbound

Fun-tuning: Characterizing the Vulnerability of Proprietary LLMs to Optimization-based Prompt Injection Attacks via the Fine-Tuning Interface cites this paper.

Fun-tuning: Characterizing the Vulnerability of Proprietary LLMs to Optimization-based Prompt Injection Attacks via the Fine-Tuning Interface PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T19:44:46.764390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:44:46.764390Z digest=sha256:74c801ed1cae081839d670ee9dde51b6da5caaa2de3ab47e2915bf88eb01d181

Observation 7f5b80dd-b813-4747-acfb-b42906214420 · inbound

KDA: A Knowledge-Distilled Attacker for Generating Diverse Prompts to Jailbreak LLMs cites this paper.

KDA: A Knowledge-Distilled Attacker for Generating Diverse Prompts to Jailbreak LLMs PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T04:22:00.571836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:22:00.571836Z digest=sha256:d0aa0b8c987707c16d149082a8466d8af558c58e178158ba871ed146ee36c846

Observation 6134f8b8-b635-44dc-a7aa-71a70786eb96 · inbound

JailbreaksOverTime: Detecting Jailbreak Attacks Under Distribution Shift cites this paper.

JailbreaksOverTime: Detecting Jailbreak Attacks Under Distribution Shift PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-16T05:57:58.326747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:57:58.326747Z digest=sha256:faa6be7f9a05956183fec661458733f4f5fe7e781ed731f2d9be8a5dc005ef3d

Observation e2e9b6ee-54bb-4ade-ad0d-4e74f7f4ed59 · inbound

Adversarial Attacks in Multimodal Systems: A Practitioner's Survey cites this paper.

Adversarial Attacks in Multimodal Systems: A Practitioner's Survey PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T00:03:02.592573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:03:02.592573Z digest=sha256:e13ee52d70e1cc5a7a076cdc8ec8d6a2f25a5b68af869f92736d5cd8ba9a36fa

Observation 5ac81ce7-660f-4fef-bd24-e4a98aaac07d · inbound

FORTRESS: Frontier Risk Evaluation for National Security and Public Safety cites this paper.

FORTRESS: Frontier Risk Evaluation for National Security and Public Safety PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:03.747220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:03.747220Z digest=sha256:c094abd910b3b0ad74b240a9a1ff08004eab1814ca81a6e755f861b2ca8f26f1

Observation 2f51017b-8333-451b-9117-773d2e2a4d40 · inbound

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models cites this paper.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:52.222527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:52.222527Z digest=sha256:a899407f48e05c0f8cf56658f16c894aebc33487e052fb58468af377e5713495

Observation b30ae07a-a659-44d6-bf7e-f39186ebaa49 · inbound

Manipulating LLM Web Agents with Indirect Prompt Injection Attack via HTML Accessibility Tree cites this paper.

Manipulating LLM Web Agents with Indirect Prompt Injection Attack via HTML Accessibility Tree PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:52:56.821563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:52:56.821563Z digest=sha256:ef13356d53fcc477852990431d23ea5ad484b6bbbcc2b848333d60579c4cb6ea

Observation a0c8bda2-c381-4fee-ac02-719defb3dbb7 · inbound

Towards terahertz nanomechanics cites this paper.

Towards terahertz nanomechanics PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T01:03:52.185970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T01:03:52.185970Z digest=sha256:e5e536a502240f190bee8004078d113c9545e2c805adf5ceb290876c5dc8d7b0

Observation 6dbccbd1-c913-4925-85b4-3920db2a1035 · inbound

ASTRA: Autonomous Spatial-Temporal Red-teaming for AI Software Assistants cites this paper.

ASTRA: Autonomous Spatial-Temporal Red-teaming for AI Software Assistants PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T01:04:36.335715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T01:04:36.335715Z digest=sha256:88609ac1380b55ec394ec4f2a27644f6d73444f36febd360c46e8f2ddeb5588a

Observation 8cd7a346-7dde-45fc-bf34-b6b32029aa32 · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:37.579293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:37.579293Z digest=sha256:f9c531dbd589089261715499458f11ebb43ecc08cb654512d204af00b9ad0ca8

Observation dc7b0358-52d1-4c7d-913a-33040afd5a44 · inbound

On Surjectivity of Neural Networks: Can you elicit any behavior from your model? cites this paper.

On Surjectivity of Neural Networks: Can you elicit any behavior from your model? PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-05T16:00:48.956894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:00:48.956894Z digest=sha256:d18c3898cd44a8ca63672dd4f8693bacc1c9a3b7a62b1f72975044d751f3f35c

Observation c527a82b-0926-4edf-b467-e9169c2ecfd7 · inbound

The Resurgence of GCG Adversarial Attacks on Large Language Models cites this paper.

The Resurgence of GCG Adversarial Attacks on Large Language Models PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T13:42:43.183386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:42:43.183386Z digest=sha256:3f219914ee8e3ad59b8650d14929728067acabb67167a500c54a97f3cd7bb068

Observation 2afdec0c-ec7b-47c6-88d9-e02135cddd6e · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 170

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:54.008721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:54.008721Z digest=sha256:ef83a4b907013f6362a39bb62428a96b4237d8e92c15c576128aa71c4afc8bde

Observation 57ec1a3d-fe0a-4f4b-b9d1-6d5e664becad · inbound

Stop Tracking Me! Proactive Defense Against Attribute Inference Attack in LLMs cites this paper.

Stop Tracking Me! Proactive Defense Against Attribute Inference Attack in LLMs PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:27:23.216512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T05:22:59.050348Z digest=sha256:0fd323f6b484e5c01f4ad602c21db155feaaea43a006c13facf3b176e03ad31d

Observation 53bc9f05-c3d8-4098-826b-763d6ad5da06 · inbound

FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption cites this paper.

FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:16:28.277270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-07T06:49:56.316472Z digest=sha256:2180b1be19cb76c39e7aa17b33b47ce967e915a8a819d807b37e32effd4551de

Observation 63eeb90d-cca1-4bae-a37f-adf7e093f5a5 · inbound

SoK: Robustness in Large Language Models against Jailbreak Attacks cites this paper.

SoK: Robustness in Large Language Models against Jailbreak Attacks PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:01:08.816756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-08T16:42:41.137808Z digest=sha256:8a3a3f6d34237f1b8bcc329c41043e1d2d9baf5d84214189a8a2a1f0f5507a64

Observation 7dae81a8-19d0-44f4-8798-af83a6e90b47 · inbound

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents cites this paper.

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:46:27.732352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-12T01:57:22.554881Z digest=sha256:800bc89886744440674d5c8aee74f7939e34ce7226ff38d6884683036ab093e2

Observation 8fe79c09-c8f7-4556-ad51-6d22acb2b52b · inbound

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents cites this paper.

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:45:45.470549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T22:52:54.684185Z digest=sha256:3752f888590f323616bb4bac94e1aedf3f2ae6bf77a68afe178598d1e7861ebc

Observation 4ba2924f-1331-42aa-b9cc-af01ca9bbc55 · inbound

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents cites this paper.

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T05:17:31.236691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:17:31.236691Z digest=sha256:e9e94e0713de70c17a70039f5b63e20d3901d4f024dea23e4e45c55b7fe72819

Observation a112aa0d-4624-4733-b074-f1a39d29c1c0 · inbound

Security in the Fine-Tuning Lifecycle of Large Language Models: Threats, Defenses,Evaluation, and Future Directions cites this paper.

Security in the Fine-Tuning Lifecycle of Large Language Models: Threats, Defenses,Evaluation, and Future Directions PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:54:03.335464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T23:51:05.413122Z digest=sha256:09bda5fe8348371fe1fd326ce258f964bc94a15de42cfbef9a4966f4ea60f48a

Observation 14368a5e-ca89-425e-9f09-f61fa9440c9d · inbound

Out of Sight: Compression-Aware Content Protection against Agentic Crawlers cites this paper.

Out of Sight: Compression-Aware Content Protection against Agentic Crawlers PAL: Proxy-Guided Black-Box Attack on Large Language Models

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-07-10T11:57:03.491182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-07-10T11:48:08.491407Z digest=sha256:de0292f71827e99bebc5203ea59cebcd4d22895a682322dd917fdf498bc394d4