Pith. sign in

Paper Citation Record · LEDGER

Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2302.05733.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2302.05733 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 24 of 24 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:26:13.657815Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

27
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 40842c4f-588b-44d0-85a7-b46df7e64d03 · inbound

Jailbroken: How Does LLM Safety Training Fail? cites this paper.

Jailbroken: How Does LLM Safety Training Fail? Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-14T18:17:42.855510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T18:17:42.752997Z digest=sha256:dae65bf1a48dda22a51a4790eebc4cd42addf3d99a08515a7db91eb66a2f61e3

Observation 890a477b-2c06-42aa-b0c7-c603ae315350 · inbound

"Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models cites this paper.

"Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-17T08:39:28.174978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T08:39:28.047394Z digest=sha256:d098c6bf4c4f86dfa97e3c9c9a7f4b87cf9376f3c162208e1872bc245ce95cb7

Observation 72287e84-52b2-4c64-83c7-763383e4bd24 · inbound

Baseline Defenses for Adversarial Attacks Against Aligned Language Models cites this paper.

Baseline Defenses for Adversarial Attacks Against Aligned Language Models Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:24:40.023866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T23:24:39.835347Z digest=sha256:d5167137943273d67924d1dd8f9ceda80290ef465914981543318e1f403941ea

Observation dec75bb5-f715-4e46-a240-8010246d6287 · inbound

AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models cites this paper.

AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-12T16:28:04.062859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T16:28:03.996446Z digest=sha256:9a9d7342a74d9876dc4f3ec1b5f3faf02fc6d8738624e2023ea78d5a527c2c4e

Observation 5ca57504-058d-4863-8fff-5f7e9942d0d9 · inbound

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation cites this paper.

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:00:51.556822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T22:00:51.487120Z digest=sha256:0115d5cb9ccc955cbffdfd60e20fbb5472318e0b9fa537866790621985c26165

Observation 8f19794a-fcc9-4122-ac96-88a49465957f · inbound

Whispers in the Machine: Confidentiality in Agentic Systems cites this paper.

Whispers in the Machine: Confidentiality in Agentic Systems Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-24T04:03:53.849912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-24T03:59:03.972043Z digest=sha256:d1613739cda04b41a1dc293eefff0824020f04371c01c3ee9ae4eff16ba8d1a7

Observation a1aa8d47-9e09-4692-b3f3-498a0ada33ba · inbound

A StrongREJECT for Empty Jailbreaks cites this paper.

A StrongREJECT for Empty Jailbreaks Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:28:02.825041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T21:28:02.745230Z digest=sha256:31e2a4cfd85f4d28f3cf36aadf87c069e85430aafa311870474957a8f906e88d

Observation 83ea2ea0-dc41-4608-8feb-36b8746ef217 · inbound

LLM Agents can Autonomously Exploit One-day Vulnerabilities cites this paper.

LLM Agents can Autonomously Exploit One-day Vulnerabilities Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:18:27.709259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T04:18:27.597704Z digest=sha256:618186fbec18c311a31dbfe140688cb11631300ab56a0ae8f009e90660620035

Observation db4aa327-8f7b-4cc0-938c-63201fa0ad10 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.691880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:7f5f4a6ca18b3a8058f6c905e42e70be1030e896e921b9fba6fc79878c6765fb

Observation d826b6f6-c434-4f2f-9f3c-ebcb4d31cbff · inbound

Prompt Infection: LLM-to-LLM Prompt Injection within Multi-Agent Systems cites this paper.

Prompt Infection: LLM-to-LLM Prompt Injection within Multi-Agent Systems Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:32:19.734280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-15T19:32:19.405615Z digest=sha256:1186a758d40473c7366ffb443c493cbf77858a188e4f57b285e8d28bf7abcd16

Observation 87dea015-967a-41c4-b047-429f37b479c2 · inbound

Faster-GCG: Efficient Discrete Optimization Jailbreak Attacks against Aligned Large Language Models cites this paper.

Faster-GCG: Efficient Discrete Optimization Jailbreak Attacks against Aligned Large Language Models Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T19:03:21.401463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T19:01:43.922848Z digest=sha256:5ede59a46f0022ec85e6c292a7e588df63def077a44d8d25739520c2175a17cc

Observation 87c3da8a-2bd1-4af8-98b4-e89550e2d730 · inbound

Circumventing Safety Alignment in Large Language Models Through Embedding Space Toxicity Attenuation cites this paper.

Circumventing Safety Alignment in Large Language Models Through Embedding Space Toxicity Attenuation Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T19:26:13.657815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:26:13.657815Z digest=sha256:64fd8e123bcb9cf509cc3fc5f71eeaecd4ef6ac2c899965368fb46d3891e8f32

Observation 503994c0-873c-4d01-bc23-7d6f96f3ef5c · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.107700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.107700Z digest=sha256:8b325590d98cbe373ee599c1f0fe4116621d68063bdfd04d4d6c6b4f23e995ba

Observation a6857368-c41a-444e-9b2e-866c399a234a · inbound

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring cites this paper.

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T14:51:03.693056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:51:03.693056Z digest=sha256:bca800c7aa330e80ffd03745c731aff0e4e8e3ab087f36a9a57e266f56b1f614

Observation 7f0c4c94-9fea-4cec-b91e-ed6b9cb57b4e · inbound

Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs cites this paper.

Breaking to Build: A Threat Model of Prompt-Based Attacks for Securing LLMs Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T05:59:27.941873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:59:27.941873Z digest=sha256:9f66be5ec4263a334b32c5b57534fd20f44b64e0aa1e861a68bb0cffba82ca69

Observation 3540126a-5c94-4b96-a78e-565f1ab7fa6f · inbound

Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint cites this paper.

Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T23:10:21.297410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:10:21.297410Z digest=sha256:847eec5b775f19d808652b6060c3361bcf5972247a88a260d863b4b037fc026c

Observation efbff674-42ba-4754-b406-ece032bb95c6 · inbound

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security cites this paper.

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T23:09:41.524212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:09:41.524212Z digest=sha256:98e0d973bde9221e6ed22ecc31999d1b546cd10acc7b913729902cedeb00e255

Observation d508c59e-88d6-4476-ab2e-39b72eba5cac · inbound

Stop Testing Attacks, Start Diagnosing Defenses: The Four-Checkpoint Framework Reveals Where LLM Safety Breaks cites this paper.

Stop Testing Attacks, Start Diagnosing Defenses: The Four-Checkpoint Framework Reveals Where LLM Safety Breaks Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T02:51:56.253042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:51:56.253042Z digest=sha256:320c22bf2b298e6e4f0712e1abcaeb2a63007e601a724593dcbfe99734cfd240

Observation 1160e7af-8cad-404a-a8de-db28bcebb7d5 · inbound

An Empirical Security Evaluation of LLM-Generated Cryptographic Rust Code cites this paper.

An Empirical Security Evaluation of LLM-Generated Cryptographic Rust Code Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:56:26.794119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T13:26:17.100399Z digest=sha256:07e04ae44679ec17d4e56d2b7b92a99f74c709b376a345657bc8002283df3a53

Observation ed14cc08-8635-43ca-add9-8f1e6044f166 · inbound

Block-wise Codeword Embedding for Reliable Multi-bit Text Watermarking cites this paper.

Block-wise Codeword Embedding for Reliable Multi-bit Text Watermarking Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:31:06.043863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-09T19:52:21.350072Z digest=sha256:e4af5181d817bf0dbe163093c7d4f3d2592a0cb6f171c1d3774dbb5f2315ff34

Observation a2877a87-885f-4c60-b65a-24eb260debe9 · inbound

Speculative Decoding at Temperature Zero: A Scoped Safety-Invariance Screen with a 48,072-Sample Expansion cites this paper.

Speculative Decoding at Temperature Zero: A Scoped Safety-Invariance Screen with a 48,072-Sample Expansion Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:59:58.296057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T00:05:06.647295Z digest=sha256:5f2a5e74b750d9687b3335e2d910fb1dd71c769259e82fc49782d0af1607b155

Observation a0791066-8569-4077-912c-6a9276d55ff4 · inbound

Security--Fidelity Tradeoffs: The Hidden Cost of Prompt Injection Defense cites this paper.

Security--Fidelity Tradeoffs: The Hidden Cost of Prompt Injection Defense Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 107

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T12:45:44.893381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-01T01:44:07.700127Z digest=sha256:e65f49a97145518e9bfbc43b275a7b7db10a82421fbaa42f080dae415bbd898e

Observation 6b1f1121-dc29-424d-8748-8d7fa42f2719 · inbound

A Lifecycle and Application-Stack Survey of Large Language Model Vulnerabilities: Attacks, Risks, Defenses, and Open Problems cites this paper.

A Lifecycle and Application-Stack Survey of Large Language Model Vulnerabilities: Attacks, Risks, Defenses, and Open Problems Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T11:05:42.347271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T04:44:23.543728Z digest=sha256:43b441a5c9d72289513eea953c1f9489ca73d964fc05e10c11b896f71f5bb7c2

Observation 9d2ee81b-6ff7-4404-86db-19c084bdfc09 · inbound

RoguePrompt: Dual-Layer Encoding for Self-Reconstruction to Circumvent LLM Moderation cites this paper.

RoguePrompt: Dual-Layer Encoding for Self-Reconstruction to Circumvent LLM Moderation Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T08:37:33.396857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:37:33.396857Z digest=sha256:e210896a987240398bb7dafed31cfe9d0505029d3d6b4101ee2e30d5081dc861