REVIEW 3 cited by
Reproducibility in Machine Learning-Driven Research
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Research is facing a reproducibility crisis, in which the results and findings of many studies are difficult or even impossible to reproduce. This is also the case in machine learning (ML) and artificial intelligence (AI) research. Often, this is the case due to unpublished data and/or source-code, and due to sensitivity to ML training conditions. Although different solutions to address this issue are discussed in the research community such as using ML platforms, the level of reproducibility in ML-driven research is not increasing substantially. Therefore, in this mini survey, we review the literature on reproducibility in ML-driven research with three main aims: (i) reflect on the current situation of ML reproducibility in various research fields, (ii) identify reproducibility issues and barriers that exist in these research fields applying ML, and (iii) identify potential drivers such as tools, practices, and interventions that support ML reproducibility. With this, we hope to contribute to decisions on the viability of different solutions for supporting ML reproducibility.
Forward citations
Cited by 3 Pith papers
-
Private, Verifiable, and Auditable AI Systems
A thesis demonstrating partial prototypes for zk-verifiable model evaluation and privacy-preserving retrieval, and arguing these pieces can compose into end-to-end auditable AI systems.
-
yProv4ML: Effortless Provenance Tracking for Machine Learning Systems
yProv4ML is a new Python library that captures ML training provenance, such as parameters, metrics, and system information, in standard PROV-JSON format with an MLFlow-like API.
-
Practical Application and Limitations of AI Certification Catalogues in the Light of the AI Act
Applying the Fraunhofer AI Assessment Catalogue to the EmoPy/RIOT emotion recognition system shows the catalogue is comprehensive but time-consuming, and that missing documentation and an inactive development team blo...
Discussion (0). Continue with ORCID to comment.