Skip to main content
Have a personal or library account? Click to login
DataDeps.jl: Repeatable Data Setup for Reproducible Data Science Cover

DataDeps.jl: Repeatable Data Setup for Reproducible Data Science

Open Access
|Oct 2019

Abstract

We present DataDeps.jl: a julia package for the reproducible handling of static datasets to enhance the repeatability of scripts used in the data and computational sciences. It is used to automate the data setup part of running software which accompanies a paper to replicate a result. This step is commonly done manually, which expends time and allows for confusion. This functionality is also useful for other packages which require data to function (e.g. a trained machine learning based model). DataDeps.jl simplifies extending research software by automatically managing the dependencies and makes it easier to run another author’s code, thus enhancing the reproducibility of data science research.

DOI: https://doi.org/10.5334/jors.244 | Journal eISSN: 2049-9647
Language: English
Submitted on: Aug 6, 2018
Accepted on: Oct 3, 2019
Published on: Oct 29, 2019
Published by: Ubiquity Press
In partnership with: Paradigm Publishing Services
Publication frequency: 1 issue per year

© 2019 Lyndon White, Roberto Togneri, Wei Liu, Mohammed Bennamoun, published by Ubiquity Press
This work is licensed under the Creative Commons Attribution 4.0 License.