Pith. sign in

Paper Citation Record · LEDGER

AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2410.09040.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.09040 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:56:34.406750Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T13:26:59.232516Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d257de41-47d1-4880-8d59-30a34f52a0b9 · inbound

Preventing Jailbreak Prompts as Malicious Tools for Cybercriminals: A Cyber Defense Perspective cites this paper.

Preventing Jailbreak Prompts as Malicious Tools for Cybercriminals: A Cyber Defense Perspective AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T12:56:34.406750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T12:56:34.406750Z digest=sha256:2feffd6e9af6c9a8218dcfe0614c29f6649e9d30d78f3481563e6272252b9cf2

Observation 7e829e40-57a1-43f8-94f4-6d394f87fdd8 · inbound

Model-Editing-Based Jailbreak against Safety-aligned Large Language Models cites this paper.

Model-Editing-Based Jailbreak against Safety-aligned Large Language Models AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T18:11:05.814369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:11:05.814369Z digest=sha256:c8e36e8f0fa54981ab8f185e4ec13db05bcd00c2531458bf5557b551d4ea44f5

Observation c9e97ec9-7496-4432-bb73-2dcc375f3398 · inbound

Benign-to-Toxic Jailbreaking: Inducing Harmful Responses from Harmless Prompts cites this paper.

Benign-to-Toxic Jailbreaking: Inducing Harmful Responses from Harmless Prompts AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T14:01:10.752332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:01:10.752332Z digest=sha256:77f0d7e725f9a05062b99e31bfd188981f8a89c2775aa4280215b664673c78f3

Observation a3456976-8d19-41d6-8adf-41f159710c7c · inbound

D-Fusion: Direct Preference Optimization for Aligning Diffusion Models with Visually Consistent Samples cites this paper.

D-Fusion: Direct Preference Optimization for Aligning Diffusion Models with Visually Consistent Samples AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T13:23:11.956259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:23:11.956259Z digest=sha256:6f47dbc994e9be9de008485fd64109719d2f88a7ca84ae7b4b34b8deaa1f50dd

Observation 56374b2a-06eb-4b3a-aa2e-88f649a29167 · inbound

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments cites this paper.

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-19T01:02:54.848700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-19T01:02:07.088724Z digest=sha256:6e5f3ce5d4592176ea781414a5ef6f3ff3e646d7268f41e135a1ee1cd8f5130a

Observation 84764f45-40d3-4ef5-b735-fd07600f2560 · inbound

On Surjectivity of Neural Networks: Can you elicit any behavior from your model? cites this paper.

On Surjectivity of Neural Networks: Can you elicit any behavior from your model? AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-05T16:00:52.007648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:00:52.007648Z digest=sha256:c371b03cf0f9fda50da82b540a2405e3f0a8e1429ca3101b5c481c29ad3cad48

Observation 9c548fb0-a63b-490d-bc83-031724e6bec0 · inbound

AgentSentinel: An End-to-End and Real-Time Security Defense Framework for Computer-Use Agents cites this paper.

AgentSentinel: An End-to-End and Real-Time Security Defense Framework for Computer-Use Agents AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T21:44:12.418917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:44:12.418917Z digest=sha256:0343fee516fb45a9476ab78715943103a67fb7a1b8d6562e818cb25f08bd6705

Observation 996ffc04-5bd4-419f-8bf2-b3f19a79be08 · inbound

Eyes-on-Me: Scalable RAG Poisoning through Transferable Attention-Steering Attractors cites this paper.

Eyes-on-Me: Scalable RAG Poisoning through Transferable Attention-Steering Attractors AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T13:28:09.231999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:28:09.231999Z digest=sha256:20381a25d9f1d055fc79239009bd4c717e4e6bcf68f0f0894e58f291c3778031

Observation 05cf7a59-9719-49e4-a2d2-8becd77062a2 · inbound

RouteHijack: Routing-Aware Attack on Mixture-of-Experts LLMs cites this paper.

RouteHijack: Routing-Aware Attack on Mixture-of-Experts LLMs AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:46:12.240917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-09T19:22:00.217729Z digest=sha256:a9051c188ab27224e146d916353403f0af44a96ad090f268bb93f79cbb2eedc9

Observation 9fb9e745-3be4-4668-b75a-7a5bc99809c9 · inbound

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks cites this paper.

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:26:59.234214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-28T01:16:07.252429Z digest=sha256:333fff82e01a2607cbd5a2d8614f98cd3e9135f851761471f2a34d001ef7c925

Observation 63302013-a04c-44d4-b23e-efff0de17d93 · inbound

Measuring the Wrong Thing: Internal Harmfulness Scores Anti-Rank Successful Jailbreaks cites this paper.

Measuring the Wrong Thing: Internal Harmfulness Scores Anti-Rank Successful Jailbreaks AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T13:57:41.302502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T13:57:41.302502Z digest=sha256:1e1eac6021b0af222baa4870ff1bf959ea8dfdf9c254c8b8555c170e3ed22e1e