REVIEW 2 cited by
The Hateful Memes Challenge: Detecting Hate Speech in Multimodal Memes
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This work proposes a new challenge set for multimodal classification, focusing on detecting hate speech in multimodal memes. It is constructed such that unimodal models struggle and only multimodal models can succeed: difficult examples ("benign confounders") are added to the dataset to make it hard to rely on unimodal signals. The task requires subtle reasoning, yet is straightforward to evaluate as a binary classification problem. We provide baseline performance numbers for unimodal models, as well as for multimodal models with various degrees of sophistication. We find that state-of-the-art methods perform poorly compared to humans (64.73% vs. 84.7% accuracy), illustrating the difficulty of the task and highlighting the challenge that this important problem poses to the community.
Forward citations
Cited by 2 Pith papers
-
On the Reliability of Vision-Language Models Under Adversarial Frequency-Domain Perturbations
Vision-language model judgments about image realism and generated captions can be shifted by imperceptible perturbations confined to specific spatial frequency bands, even under black-box access.
-
The Ethics of Generative AI in Anonymous Spaces: A Case Study of 4chan's /pol/ Board
A case study of 66 AI-generated images from 4chan's /pol/ board finds 28.8% contain racist content and 28.8% anti-Semitic content, but the sample is small and likely skewed.
Discussion (0). Continue with ORCID to comment.