Pith. sign in

REVIEW 1 cited by

Investigating Labeler Bias in Face Annotation for Machine Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2301.09902 v3 pith:72FGDGQZ submitted 2023-01-24 cs.LG cs.HC

classification cs.LGcs.HC
keywords labelerbiasartificialintelligencedatasetsprocesssubsequentlytraining
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In a world increasingly reliant on artificial intelligence, it is more important than ever to consider the ethical implications of artificial intelligence on humanity. One key under-explored challenge is labeler bias, which can create inherently biased datasets for training and subsequently lead to inaccurate or unfair decisions in healthcare, employment, education, and law enforcement. Hence, we conducted a study to investigate and measure the existence of labeler bias using images of people from different ethnicities and sexes in a labeling task. Our results show that participants possess stereotypes that influence their decision-making process and that labeler demographics impact assigned labels. We also discuss how labeler bias influences datasets and, subsequently, the models trained on them. Overall, a high degree of transparency must be maintained throughout the entire artificial intelligence training process to identify and correct biases in the data as early as possible.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Stress-Testing ML Pipelines with Adversarial Data Corruption

    cs.LG 2025-06 conditional novelty 7.0 of 10

    SAVAGE uses dependency graphs plus beam search and Bayesian optimization to find structured data corruptions that degrade ML pipelines far more than random or manual errors.

Pith tools