Pith. sign in

Paper Citation Record · LEDGER

Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2404.19287.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.19287 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:04:01.363663Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T13:23:28.107852Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c45f3868-043e-4b3e-98c7-f58623422f7e · inbound

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety cites this paper.

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 243

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:42:34.230016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T04:39:04.591722Z digest=sha256:581cf39f02fde47eea779e622e9d9168786bb2f27f2e531ef957ae87fec9b095

Observation b9b58612-b4bb-4b69-a49d-d0ceda8543f4 · inbound

BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization cites this paper.

BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:01.363663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:01.363663Z digest=sha256:921dc6cc7ca34a76b88f3d45baed2e5d6c92136e4e647064bbc98990907b0429

Observation ed9c31f9-b857-487a-bd06-383e5ae4aa40 · inbound

Dual-Path Stable Soft Prompt Generation for Domain Generalization cites this paper.

Dual-Path Stable Soft Prompt Generation for Domain Generalization Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:34.741589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:34.741589Z digest=sha256:c944c871f73d68894eff4644aa34343a344df0f99843b3864e9b0ce53bc1989d

Observation 48fb0b46-4211-47e3-ade7-8dedcfe9dfa8 · inbound

Coordinated Robustness Evaluation Framework for Vision-Language Models cites this paper.

Coordinated Robustness Evaluation Framework for Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T10:40:49.610244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:40:49.610244Z digest=sha256:8027627f200f664e227e353a914815dba635f1d64210c16fedb4f06ad102213e

Observation 6cdb1c89-ffa7-413e-98b4-26347ee70ee5 · inbound

On the Feasibility of Poisoning Text-to-Image AI Models via Adversarial Mislabeling cites this paper.

On the Feasibility of Poisoning Text-to-Image AI Models via Adversarial Mislabeling Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:50.267956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:50.267956Z digest=sha256:3059579459a623504225f0915951ab26d6f7ccbf5af098198ea2f360d9d00eba

Observation d3863703-51e5-41df-9d6c-1d1c333341cb · inbound

Invisible Injections: Exploiting Vision-Language Models Through Steganographic Prompt Embedding cites this paper.

Invisible Injections: Exploiting Vision-Language Models Through Steganographic Prompt Embedding Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T11:54:43.081752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:54:43.081752Z digest=sha256:20b57ea26652e45faea084d1ea09bcf62eed100d7765cb6d9f2fd328893e2677

Observation 460e04dd-8b0b-4288-91bb-5d1f87039018 · inbound

ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models cites this paper.

ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-18T15:26:33.883060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T15:23:13.318310Z digest=sha256:c5e369be8f3ef7a42fc718106bbc17df8a8e44d05740c5e3c6ae5d11f57910b7

Observation 7178a73d-70d6-497f-a870-60ef53d4a74a · inbound

ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models cites this paper.

ORCA: An Agentic Reasoning Framework for Hallucination and Adversarial Robustness in Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-21T21:30:39.418151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T21:28:59.898029Z digest=sha256:a94580938876b2cd4b89160f8d826188fef6fd028dcffc34289f8f48ff764652

Observation 25df41f2-ed47-49fe-b0ad-c51c755b6b34 · inbound

Contrastive Spectral Rectification: Test-Time Defense towards Zero-shot Adversarial Robustness of CLIP cites this paper.

Contrastive Spectral Rectification: Test-Time Defense towards Zero-shot Adversarial Robustness of CLIP Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T07:47:37.588939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T07:47:37.588939Z digest=sha256:2e6ffb93731eff3523392003c7e25faaf10f308cc18b7ec18c9df713a08567b7

Observation 537782d0-b461-44ba-a0c9-ee9e00e5ff48 · inbound

Revealing Physical-World Semantic Vulnerabilities: Universal Adversarial Patch for Infrared Vision-Language Models cites this paper.

Revealing Physical-World Semantic Vulnerabilities: Universal Adversarial Patch for Infrared Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T19:48:11.625968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T19:43:29.058335Z digest=sha256:1fe80255d84adeb2eeea1367bd8f3b3ddec3cd37642826f795aaea375cb56c4d

Observation 361705f3-d9fa-4e77-a6f4-173ccdcbdf28 · inbound

Revealing Physical-World Semantic Vulnerabilities: Universal Adversarial Patch for Infrared Vision-Language Models cites this paper.

Revealing Physical-World Semantic Vulnerabilities: Universal Adversarial Patch for Infrared Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T16:52:02.261328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:52:02.261328Z digest=sha256:27e8bca4cdd7e2fbf5a12ce3beacd8f38b02ae673a443bcffcf09a7162d50b51

Observation ca664a2a-660c-4bd0-8455-5fd486f25e56 · inbound

Hierarchical Attacks for Multi-Modal Multi-Agent Reasoning cites this paper.

Hierarchical Attacks for Multi-Modal Multi-Agent Reasoning Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:19:27.931230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T20:16:30.070819Z digest=sha256:92501bf4f95f624063eb1aa2060bbd595118e2e704aa71f74f23456fc10dea53

Observation ef32f8f9-1cf9-435e-9468-f114b735f816 · inbound

TAME: Test-Time Adversarial Prompt Tuning via Mixture-of-Experts for Vision-Language Models cites this paper.

TAME: Test-Time Adversarial Prompt Tuning via Mixture-of-Experts for Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:23:21.496636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T14:20:49.278545Z digest=sha256:87ea9919fc9747697623882f71b4c4e8f314f65059efab11b140c53fedfa4d87

Observation 9e9b3f9d-0f8a-47cf-ae78-e3f053369026 · inbound

JECA^2: Judgment-Explanation Consistent Adversarial Attack against Forensic Vision-Language Models cites this paper.

JECA^2: Judgment-Explanation Consistent Adversarial Attack against Forensic Vision-Language Models Revisiting the Adversarial Robustness of Vision Language Models: a Multimodal Perspective

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T13:23:28.109593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T13:18:46.414179Z digest=sha256:e59161c0cc0398dcab43fa5764a9dd2bd2f671aba4ad99e5ed25cb19cbe44293