An audit of the Swiss Judgment Prediction Dataset finds that words labeled socially biased mostly reflect neutral legal language and can mislead bias measurements in legal AI.
Identifying biases in legal data: An algorithmic fairness perspective
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
The need to address representation biases and sentencing disparities in legal case data has long been recognized. Here, we study the problem of identifying and measuring biases in large-scale legal case data from an algorithmic fairness perspective. Our approach utilizes two regression models: A baseline that represents the decisions of a "typical" judge as given by the data and a "fair" judge that applies one of three fairness concepts. Comparing the decisions of the "typical" judge and the "fair" judge allows for quantifying biases across demographic groups, as we demonstrate in four case studies on criminal data from Cook County (Illinois).
citation-role summary
citation-polarity summary
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Analyzing Bias in Swiss Federal Supreme Court Judgments Using Facebook's Holistic Bias Dataset: Implications for Language Model Training
An audit of the Swiss Judgment Prediction Dataset finds that words labeled socially biased mostly reflect neutral legal language and can mislead bias measurements in legal AI.