Pith. sign in

REVIEW 3 cited by

DIG In: Evaluating Disparities in Image Generations with Indicators for Geographic Diversity

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2308.06198 v3 pith:ZODOXJSO submitted 2023-08-11 cs.CV cs.HC

classification cs.CVcs.HC
keywords geographicindicatorssystemscontentcreationdiversitydisparitiesvisual
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The unprecedented photorealistic results achieved by recent text-to-image generative systems and their increasing use as plug-and-play content creation solutions make it crucial to understand their potential biases. In this work, we introduce three indicators to evaluate the realism, diversity and prompt-generation consistency of text-to-image generative systems when prompted to generate objects from across the world. Our indicators complement qualitative analysis of the broader impact of such systems by enabling automatic and efficient benchmarking of geographic disparities, an important step towards building responsible visual content creation systems. We use our proposed indicators to analyze potential geographic biases in state-of-the-art visual content creation systems and find that: (1) models have less realism and diversity of generations when prompting for Africa and West Asia than Europe, (2) prompting with geographic information comes at a cost to prompt-consistency and diversity of generated images, and (3) models exhibit more region-level disparities for some objects than others. Perhaps most interestingly, our indicators suggest that progress in image generation quality has come at the cost of real-world geographic representation. Our comprehensive evaluation constitutes a crucial step towards ensuring a positive experience of visual content creation for everyone.

Discussion (0). Sign in to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Adultification Bias in LLMs and Text-to-Image Models

    cs.CY 2025-06 conditional novelty 6.0 of 10

    Large language and text-to-image models show measurable adultification bias, portraying Black girls as more mature, culpable, and sexualized than White girls in several tested models.

  2. Evaluation of Cultural Competence of Vision-Language Models

    cs.CV 2025-05 conditional novelty 6.0 of 10

    The paper proposes five theory-informed frameworks from visual cultural studies for evaluating cultural competence in vision-language models.

  3. VideoGuard: Protecting Video Content from Unauthorized Editing

    cs.CV 2025-08 unverdicted novelty 5.0 of 10

    VideoGuard adds joint, motion-aware perturbations to videos to block unauthorized diffusion-model editing.

Pith tools