REVIEW 7 cited by
A Survey on Bias and Fairness in Machine Learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
With the widespread use of AI systems and applications in our everyday lives, it is important to take fairness issues into consideration while designing and engineering these types of systems. Such systems can be used in many sensitive environments to make important and life-changing decisions; thus, it is crucial to ensure that the decisions do not reflect discriminatory behavior toward certain groups or populations. We have recently seen work in machine learning, natural language processing, and deep learning that addresses such challenges in different subdomains. With the commercialization of these systems, researchers are becoming aware of the biases that these applications can contain and have attempted to address them. In this survey we investigated different real-world applications that have shown biases in various ways, and we listed different sources of biases that can affect AI applications. We then created a taxonomy for fairness definitions that machine learning researchers have defined in order to avoid the existing bias in AI systems. In addition to that, we examined different domains and subdomains in AI showing what researchers have observed with regard to unfair outcomes in the state-of-the-art methods and how they have tried to address them. There are still many future directions and solutions that can be taken to mitigate the problem of bias in AI systems. We are hoping that this survey will motivate researchers to tackle these issues in the near future by observing existing work in their respective fields.
Forward citations
Cited by 7 Pith papers
-
Prototypicality Bias Reveals Blindspots in Multimodal Evaluation Metrics
Prototypicality bias: common text-to-image metrics systematically prefer plausible-but-wrong images over correct non-prototypical ones; PROTOSCORE mitigates but does not eliminate the failure.
-
A Theory of Inference Compute Scaling: Reasoning through Directed Stochastic Skill Search
A skill-graph random-walk model gives closed-form accuracy-versus-compute formulas for four reasoning strategies and connects them to training scaling.
-
Safe and Certifiable AI Systems: Concepts, Challenges, and Lessons Learned
The paper presents the TÜV AUSTRIA Trusted AI audit catalog, a statistical framework based on the Stochastic Application Domain Definition, minimum performance requirements, and independent-sample testing for certifyi...
-
Diversity and Inclusion in AI: Insights from a Survey of AI/ML Practitioners
A survey of 61 AI/ML practitioners finds that while most believe diverse teams and data reduce bias, actual practices like bias audits and post-development D&I checks are inconsistent and often missing.
-
How to Evaluate Automatic Speech Recognition: Comparing Different Performance and Bias Measures
On Dutch end-to-end ASR systems, average word error rate hides large performance gaps across speaker groups, and bias mitigation can lower the average while increasing bias.
-
Chatbot Deployment Considerations for Application-Agnostic Human-Machine Dialogues
Microsoft's Tay chatbot failed after learning offensive Twitter content within 16 hours, and the paper draws deployment lessons from that incident.
-
The Role of AI in Early Detection of Life-Threatening Diseases: A Retinal Imaging Perspective
This is a narrative review of AI-enhanced retinal imaging for detecting systemic diseases; it presents no new data and is undermined by multiple citation errors.
Discussion (0). Continue with ORCID to comment.