Pith. sign in

REVIEW 4 cited by

ViSAGe: A Global-Scale Analysis of Visual Stereotypes in Text-to-Image Generation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2401.06310 v3 pith:2W4H745H submitted 2024-01-12 cs.CV cs.CLcs.CY

classification cs.CVcs.CLcs.CY
keywords stereotypesvisagedepictionsgroupsidentitystereotypicalvisualpull
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Recent studies have shown that Text-to-Image (T2I) model generations can reflect social stereotypes present in the real world. However, existing approaches for evaluating stereotypes have a noticeable lack of coverage of global identity groups and their associated stereotypes. To address this gap, we introduce the ViSAGe (Visual Stereotypes Around the Globe) dataset to enable the evaluation of known nationality-based stereotypes in T2I models, across 135 nationalities. We enrich an existing textual stereotype resource by distinguishing between stereotypical associations that are more likely to have visual depictions, such as `sombrero', from those that are less visually concrete, such as 'attractive'. We demonstrate ViSAGe's utility through a multi-faceted evaluation of T2I generations. First, we show that stereotypical attributes in ViSAGe are thrice as likely to be present in generated images of corresponding identities as compared to other attributes, and that the offensiveness of these depictions is especially higher for identities from Africa, South America, and South East Asia. Second, we assess the stereotypical pull of visual depictions of identity groups, which reveals how the 'default' representations of all identity groups in ViSAGe have a pull towards stereotypical depictions, and that this pull is even more prominent for identity groups from the Global South. CONTENT WARNING: Some examples contain offensive stereotypes.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Customize Multi-modal RAI Guardrails with Precedent-based predictions

    cs.LG 2025-07 conditional novelty 6.0 of 10

    Conditioning a multimodal guardrail on retrieved 'precedent' reasoning traces, rather than static policy definitions, improves few-shot and novel-policy content-moderation F1 scores on UnsafeBench.

  2. AI's Blind Spots: Geographic Knowledge and Diversity Deficit in Generated Urban Scenario

    cs.AI 2025-06 conditional novelty 5.0 of 10

    Diffusion models FLUX 1 and SD 3.5 encode fine-grained US geographic knowledge when prompted with states or capitals, but the generic prompt 'USA' produces a metropolitan stereotype that under-represents rural, fronti...

  3. VBench++: Comprehensive and Versatile Benchmark Suite for Video Generative Models

    cs.CV 2024-11 conditional novelty 5.0 of 10

    VBench++ is a benchmark that scores text-to-video and image-to-video models on 16 quality dimensions plus trustworthiness, reporting human-alignment correlations for each.

  4. Evaluating Generative AI Systems is a Social Science Measurement Challenge

    cs.CY 2024-11 conditional novelty 4.0 of 10

    The paper argues that generative AI evaluation should adopt a social science measurement framework with an explicit 'systematized concept' between high-level ideas and concrete instruments.

Pith tools