Pith. sign in

Paper Citation Record · LEDGER

UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2405.03486.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.03486 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:50:44.351448Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:40:06.590389Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 460b2d94-3017-4fd7-bf95-0fceb5d50742 · inbound

VModA: An Effective Framework for Adaptive NSFW Image Moderation cites this paper.

VModA: An Effective Framework for Adaptive NSFW Image Moderation UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:50:44.351448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:50:44.351448Z digest=sha256:fef6dce6fb2640152ad5959834570a6ead29cadc197570f21bb5f770a64fabf8

Observation 1aa2458a-4eb7-4ac5-8c6b-f652f18bcc17 · inbound

Bridging the Gap in Vision Language Models in Identifying Unsafe Concepts Across Modalities cites this paper.

Bridging the Gap in Vision Language Models in Identifying Unsafe Concepts Across Modalities UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T17:21:35.300886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:21:35.300886Z digest=sha256:22a559ef4820a7d8b9b6ff02df661f4cfadc1712f3b1a209f34581ab4e17510f

Observation 486054fb-1c29-405f-a1e9-74cbe132e51b · inbound

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation cites this paper.

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T12:54:36.257427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:54:36.257427Z digest=sha256:4160cb9488050d1107e4ef017fc484aafe644b1079ed352cd887ccecfad7089d

Observation c6a3f856-3ca6-446d-b3c3-62ed75f6db10 · inbound

Hate in Plain Sight: On the Risks of Moderating AI-Generated Hateful Illusions cites this paper.

Hate in Plain Sight: On the Risks of Moderating AI-Generated Hateful Illusions UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T11:34:27.797808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:34:27.797808Z digest=sha256:812efc833f002893f5cab96667f798af7fe13e16b41fad1e7e4ff3ed4201ceba

Observation 49111154-6b9d-49eb-918d-f8b808234b95 · inbound

Immunizing Images from Text to Image Editing via Adversarial Cross-Attention cites this paper.

Immunizing Images from Text to Image Editing via Adversarial Cross-Attention UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T17:56:30.909485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:56:30.909485Z digest=sha256:b90ca9b42bed36fbd29920f3a994ec4c5f3d403006e8619d390ffeb25ef68a5c

Observation a781b074-3156-48d9-8cff-35e35bdff06a · inbound

SPQR: A Multi-Dimensional Benchmark for Safety Alignment under Benign Model Adaptation cites this paper.

SPQR: A Multi-Dimensional Benchmark for Safety Alignment under Benign Model Adaptation UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T20:39:57.598200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:39:57.598200Z digest=sha256:c09ed09e5637e8b7e7a6a61e1d71f1c659d9918773fd50b29c50d16468e31d22

Observation 64fd0f11-ae1e-4fed-92d1-d974cac61525 · inbound

SafeGen-Bench: Benchmarking Safety in Image-Conditioned Text-to-Video Generation cites this paper.

SafeGen-Bench: Benchmarking Safety in Image-Conditioned Text-to-Video Generation UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T17:12:25.191446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T17:05:57.685728Z digest=sha256:8e605f0394a535d799d03503c9639e038b73c582bb884c5fc6db23eb371977dc

Observation d5b7a14f-98c6-4268-aff7-72963b4f508a · inbound

VPA-Guard: Defending and Benchmarking Image-to-Video Generation Against Visual Prompt Attacks cites this paper.

VPA-Guard: Defending and Benchmarking Image-to-Video Generation Against Visual Prompt Attacks UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:40:06.591963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-25T21:10:28.989297Z digest=sha256:3596a272ec07489f78a80912535938695880b665075e58b5e3290104411cd742