Pith. sign in

Paper Citation Record · LEDGER

Safety of Multimodal Large Language Models on Images and Texts

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2402.00357.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.00357 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T15:38:17.103048Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T01:27:30.991837Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fd9d8eb5-393c-4e58-a585-61fa9825b70d · inbound

When Data Manipulation Meets Attack Goals: An In-depth Survey of Attacks for VLMs cites this paper.

When Data Manipulation Meets Attack Goals: An In-depth Survey of Attacks for VLMs Safety of Multimodal Large Language Models on Images and Texts

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T15:38:17.103048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:38:17.103048Z digest=sha256:e4b50d1b593402a8f36b610b80e875fde40ad1c0d03a8e194d19bf8ac136cc4b

Observation f71db071-708a-45be-900d-a66b1762284c · inbound

A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations cites this paper.

A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations Safety of Multimodal Large Language Models on Images and Texts

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T19:45:18.435520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T19:45:18.435520Z digest=sha256:569b24c2f07841f2cafd1e0c889bdf277fa1746cfb801a5367dd5ac5aa756e8c

Observation cccdc95a-89aa-4648-9e47-aaf4bb2cc0e4 · inbound

Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment cites this paper.

Adversarial Attacks against Closed-Source MLLMs via Feature Optimal Alignment Safety of Multimodal Large Language Models on Images and Texts

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:35:53.646908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:35:53.646908Z digest=sha256:739a539f7456cad3c2dab4b7c59359c83c34e2dd28dcb061fcc76ec82021c4d4

Observation d2bbf271-6201-44be-a588-6223836e8b4e · inbound

The First Differentiable Transfer-Based Algorithm for Discrete MicroLED Repair cites this paper.

The First Differentiable Transfer-Based Algorithm for Discrete MicroLED Repair Safety of Multimodal Large Language Models on Images and Texts

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T22:21:02.805074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:21:02.805074Z digest=sha256:da1b26d8388e8177646ec5e2735327d8ecf4fc17afb941fd4d03cc46de0498cb

Observation cbad9768-587a-4bb5-8ba6-3e50a50fa5e4 · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Safety of Multimodal Large Language Models on Images and Texts

Reference 146

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:46.157970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:46.157970Z digest=sha256:5f2b48467a1a42582b4ebb66e9613fc30ff0d4f8db7b69d0c8a006dc751fc357

Observation 3894a040-3054-4e18-a724-87ac4b6121f9 · inbound

Is GPT-4o mini Blinded by its Own Safety Filters? Exposing the Multimodal-to-Unimodal Bottleneck in Hate Speech Detection cites this paper.

Is GPT-4o mini Blinded by its Own Safety Filters? Exposing the Multimodal-to-Unimodal Bottleneck in Hate Speech Detection Safety of Multimodal Large Language Models on Images and Texts

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T16:30:29.384240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:30:29.384240Z digest=sha256:aee9d01cb5ec5ad45adfe3b5cf1f6fd2be3140512a84fcc0b2574f927a90b671

Observation 97a40a27-a51b-49e8-b125-c291234dd07f · inbound

Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing cites this paper.

Guaranteed Jailbreaking Defense via Disrupt-and-Rectify Smoothing Safety of Multimodal Large Language Models on Images and Texts

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:27.469126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T04:50:08.866969Z digest=sha256:c670455a75074b338428612a8902d9d65c7498041d6cf9c6aee622b3236240f1

Observation a3431651-ae9c-4ad6-9d92-b0bb7f6ad56a · inbound

Investigating Adversarial Robustness of Multi-modal Large Language Models cites this paper.

Investigating Adversarial Robustness of Multi-modal Large Language Models Safety of Multimodal Large Language Models on Images and Texts

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:06:27.639652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T11:11:34.152223Z digest=sha256:de2141119d9ba15afed4b000aefc13d339cfa8934587272569fb82ef7d4c2f76

Observation 339e86b5-919c-4603-ab5f-e6107eca8397 · inbound

Unveiling Privacy Risks in Multi-modal Large Language Models: Task-specific Vulnerabilities and Mitigation Challenges cites this paper.

Unveiling Privacy Risks in Multi-modal Large Language Models: Task-specific Vulnerabilities and Mitigation Challenges Safety of Multimodal Large Language Models on Images and Texts

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:27:30.993178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T16:33:28.848573Z digest=sha256:420e06482a03cd712413d2ab73529ed0777022aff48d652930737e2cd8baf155

Observation d54a7677-0c02-45c0-a740-7977675c926f · inbound

V-DEAL: Diagnosing Video Safety De-Calibration as an Understanding-Refusal Coupling Failure cites this paper.

V-DEAL: Diagnosing Video Safety De-Calibration as an Understanding-Refusal Coupling Failure Safety of Multimodal Large Language Models on Images and Texts

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T08:23:25.032757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:23:25.032757Z digest=sha256:5cf8bd3a3ab67875e3e5a9c07fafba49c01e900f5fcee74067e1d86621a57dfb