Ten popular Stable Diffusion models generated harmful images for most test prompts, showed almost no refusal behavior, and displayed a bias associating Black individuals with violence.
safe” when they are not detected by any of our classifier’s category, and “unsafe
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CY 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
When Image Generation Goes Wrong: A Safety Analysis of Stable Diffusion Models
Ten popular Stable Diffusion models generated harmful images for most test prompts, showed almost no refusal behavior, and displayed a bias associating Black individuals with violence.