Pith. sign in

Paper Citation Record · LEDGER

Are aligned neural networks adversarially aligned?

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2306.15447.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.15447 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:45:42.853648Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

25
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 70b52c00-5d1e-4019-8c6b-dbede855f858 · inbound

Universal and Transferable Adversarial Attacks on Aligned Language Models cites this paper.

Universal and Transferable Adversarial Attacks on Aligned Language Models Are aligned neural networks adversarially aligned?

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-24T07:44:08.583092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-24T07:42:09.112946Z digest=sha256:610be1d6ad1e110fe60da4d29e9b127f988161ff1fcb7eed064b75bf428e503a

Observation 85a81cbd-593d-4fd0-b8d4-ab1c37255d06 · inbound

Baseline Defenses for Adversarial Attacks Against Aligned Language Models cites this paper.

Baseline Defenses for Adversarial Attacks Against Aligned Language Models Are aligned neural networks adversarially aligned?

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:24:40.058993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T23:24:39.835347Z digest=sha256:b3f5d2ecfd9425f9adb7e011ebe7c1a92168e65d749c0e3f1339ff992c96e1a7

Observation aaa59a1f-99e0-40dd-96fc-eb6c8f354cf4 · inbound

Low-Resource Languages Jailbreak GPT-4 cites this paper.

Low-Resource Languages Jailbreak GPT-4 Are aligned neural networks adversarially aligned?

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T09:24:14.078179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:f6caad9cc8c3cbacffc0a2956d3c0dc524e5fd368be2e3b1880c60c2c574ac5d

Observation aac2d768-ee6f-446e-9d45-521ced7a6b5e · inbound

SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks cites this paper.

SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks Are aligned neural networks adversarially aligned?

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:11:00.979943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T17:11:00.639293Z digest=sha256:1c7e030bb6ef67281a07998e6f8b56ed0359b962a13ad05c7afb83eef5140f59

Observation 3403e9d2-551a-403b-8b32-e16d0ad468b5 · inbound

Improved Baselines with Visual Instruction Tuning cites this paper.

Improved Baselines with Visual Instruction Tuning Are aligned neural networks adversarially aligned?

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T19:11:34.013332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T19:11:33.783746Z digest=sha256:ff9fa498f8d766f582ee2aa421a41ae7448af56a19a7f26b9e88d6c3031ab219

Observation 98085e8b-d96b-4ffd-a0be-7156b031366a · inbound

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation cites this paper.

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation Are aligned neural networks adversarially aligned?

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T22:00:51.514305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T22:00:51.487120Z digest=sha256:8cb404d401ab3f6a83c3c711bba81a70c15c15fa655dfd797f1beccd146fe536

Observation 8ddb4cf7-b71d-4e9a-bb31-fda093a272d0 · inbound

Scalable Extraction of Training Data from (Production) Language Models cites this paper.

Scalable Extraction of Training Data from (Production) Language Models Are aligned neural networks adversarially aligned?

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T19:00:46.800869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T19:00:46.708242Z digest=sha256:1ba77241dda90fc5af19b92b0085da7bbb7b7d01bf84afd2e446b4141f4bbf88

Observation df9b70f8-16a7-438d-89c7-9cbe07c6315d · inbound

Adversarial Hubness in Multi-Modal Retrieval cites this paper.

Adversarial Hubness in Multi-Modal Retrieval Are aligned neural networks adversarially aligned?

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T06:42:39.688561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-23T06:39:36.039613Z digest=sha256:7d9f4a1a2495e518023ab5375e5d3371cf5abf5cb681710ed1ae3ebd57995195

Observation f9454b0b-8c8d-44ad-898e-cd51dce5e8ea · inbound

Understanding the Supply Chain and Risks of Large Language Model Applications cites this paper.

Understanding the Supply Chain and Risks of Large Language Model Applications Are aligned neural networks adversarially aligned?

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:42.853648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:45:42.853648Z digest=sha256:980c3216f060e190f2b9db91e738c5dc77738bf84e605c60c62c284678b8aea6

Observation 654a4353-6e73-416b-9b94-a353778e90be · inbound

Invitation Is All You Need! Promptware Attacks Against LLM-Powered Assistants in Production Are Practical and Dangerous cites this paper.

Invitation Is All You Need! Promptware Attacks Against LLM-Powered Assistants in Production Are Practical and Dangerous Are aligned neural networks adversarially aligned?

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T19:39:09.544533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:39:09.544533Z digest=sha256:9b1a74730ce6701dd1c1b1017d2827c3bfdfc105b16953df6b2277e1ada0ffdc

Observation 4ad71a2e-fd11-4cfe-b04f-9a463724be41 · inbound

AttackEval: A Systematic Empirical Study of Prompt Injection Attack Effectiveness Against Large Language Models cites this paper.

AttackEval: A Systematic Empirical Study of Prompt Injection Attack Effectiveness Against Large Language Models Are aligned neural networks adversarially aligned?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-13T12:52:49.101462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:52:49.101462Z digest=sha256:eeb1f630923c84a80a82fdbf4d0026d6b5534f16e90575c88a451d56df3dd123

Observation 308e3d43-e2c8-4ec0-be1e-e0507397f344 · inbound

VisInject: Disruption != Injection -- A Dual-Dimension Evaluation of Universal Adversarial Attacks on Vision-Language Models cites this paper.

VisInject: Disruption != Injection -- A Dual-Dimension Evaluation of Universal Adversarial Attacks on Vision-Language Models Are aligned neural networks adversarially aligned?

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:56:08.052543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-09T14:24:48.999632Z digest=sha256:75497b934921eeec4d9603a0569671726569a0f7480269893b21575ade789a56

Observation f2e3a91d-658d-45fc-b9f6-6ec9edc6e50c · inbound

Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours cites this paper.

Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours Are aligned neural networks adversarially aligned?

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:51:47.107368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-07T16:06:18.057868Z digest=sha256:00bfbf338762a1f25191fec855a26f03d4517362623ba5915b259673d0809fde

Observation 10a509cc-6079-4212-9bb1-f9b74c54e56a · inbound

Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing cites this paper.

Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing Are aligned neural networks adversarially aligned?

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:51:27.599788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:50:08.866969Z digest=sha256:2205ba24393310d88ffee7eeae7c13966d25ed5e3dbe9c7b80084e723e974171

Observation ee06d903-7470-4563-a7eb-d91c51946392 · inbound

Attention Hijacking: Response Manipulation Across Queries in Vision-Language Models cites this paper.

Attention Hijacking: Response Manipulation Across Queries in Vision-Language Models Are aligned neural networks adversarially aligned?

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:28:21.591740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T14:25:10.603780Z digest=sha256:65980013d29e242303fa7eef3d2d6118f839851b034beb232b6b41e9016ff76f

Observation 532e644e-4b6d-4b0d-90e2-f46eb8df4150 · inbound

SCI-Defense: Defending Manipulation Attacks from Generative Engine Optimization cites this paper.

SCI-Defense: Defending Manipulation Attacks from Generative Engine Optimization Are aligned neural networks adversarially aligned?

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T07:44:42.727058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T07:43:00.837852Z digest=sha256:fc04cbf79fb71f7cbf5239010a04d3c676b160c099e96a84153a3662d586a4bd

Observation b061b0a4-61be-4f8d-8f3a-99216c013c0c · inbound

RecurGuard: Runtime Monitoring for Reasoning-Token Consumption Attacks cites this paper.

RecurGuard: Runtime Monitoring for Reasoning-Token Consumption Attacks Are aligned neural networks adversarially aligned?

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:27:24.527682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T19:45:30.671490Z digest=sha256:a0b0ac119c40e328128bafd12ce507aa62d42c34e9bcaa0d51d5f7efd6b18f84