Pith. sign in

REVIEW 3 cited by

Reproducibility in Machine Learning-Driven Research

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2307.10320 v1 pith:IB4S2XQZ submitted 2023-07-19 cs.LG cs.CYstat.ME

classification cs.LGcs.CYstat.ME
keywords reproducibilityresearchcasedifferentfieldsidentifymachineml-driven
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Research is facing a reproducibility crisis, in which the results and findings of many studies are difficult or even impossible to reproduce. This is also the case in machine learning (ML) and artificial intelligence (AI) research. Often, this is the case due to unpublished data and/or source-code, and due to sensitivity to ML training conditions. Although different solutions to address this issue are discussed in the research community such as using ML platforms, the level of reproducibility in ML-driven research is not increasing substantially. Therefore, in this mini survey, we review the literature on reproducibility in ML-driven research with three main aims: (i) reflect on the current situation of ML reproducibility in various research fields, (ii) identify reproducibility issues and barriers that exist in these research fields applying ML, and (iii) identify potential drivers such as tools, practices, and interventions that support ML reproducibility. With this, we hope to contribute to decisions on the viability of different solutions for supporting ML reproducibility.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Private, Verifiable, and Auditable AI Systems

    cs.CR 2025-08 conditional novelty 4.0 of 10

    A thesis demonstrating partial prototypes for zk-verifiable model evaluation and privacy-preserving retrieval, and arguing these pieces can compose into end-to-end auditable AI systems.

  2. yProv4ML: Effortless Provenance Tracking for Machine Learning Systems

    cs.LG 2025-07 conditional novelty 4.0 of 10

    yProv4ML is a new Python library that captures ML training provenance, such as parameters, metrics, and system information, in standard PROV-JSON format with an MLFlow-like API.

  3. Practical Application and Limitations of AI Certification Catalogues in the Light of the AI Act

    cs.CY 2025-01 conditional novelty 4.0 of 10

    Applying the Fraunhofer AI Assessment Catalogue to the EmoPy/RIOT emotion recognition system shows the catalogue is comprehensive but time-consuming, and that missing documentation and an inactive development team blo...

Pith tools