REVIEW 2 cited by
Tools and Practices for Responsible AI Engineering
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Responsible Artificial Intelligence (AI) - the practice of developing, evaluating, and maintaining accurate AI systems that also exhibit essential properties such as robustness and explainability - represents a multifaceted challenge that often stretches standard machine learning tooling, frameworks, and testing methods beyond their limits. In this paper, we present two new software libraries - hydra-zen and the rAI-toolbox - that address critical needs for responsible AI engineering. hydra-zen dramatically simplifies the process of making complex AI applications configurable, and their behaviors reproducible. The rAI-toolbox is designed to enable methods for evaluating and enhancing the robustness of AI-models in a way that is scalable and that composes naturally with other popular ML frameworks. We describe the design principles and methodologies that make these tools effective, including the use of property-based testing to bolster the reliability of the tools themselves. Finally, we demonstrate the composability and flexibility of the tools by showing how various use cases from adversarial robustness and explainable AI can be concisely implemented with familiar APIs.
Forward citations
Cited by 2 Pith papers
-
An Empirical Study on Decision-Making Aspects in Responsible Software Engineering for AI
Practitioners in AI software development report that ethical guidelines are rarely operationalized, and decision-making is driven more by personal values and organizational culture than by formal frameworks.
-
Robust Training with Data Augmentation for Medical Imaging Classification
A one-line change to RobustAugMix, applying cross-entropy loss to adversarial examples, is benchmarked on three medical imaging datasets.
Discussion (0). Continue with ORCID to comment.