Computational reproducibility of Jupyter notebooks from biomedical publications

Sheeba Samuel; Daniel Mietchen

Journal ArticleOPEN ACCESS

Computational reproducibility of Jupyter notebooks from biomedical publications

GigaScience (2024) 13

DOI: 10.1093/gigascience/giad113

4Citations

34Readers

Abstract

Background: Jupyter notebooks facilitate the bundling of executable code with its documentation and output in one interactive environment, and they represent a popular mechanism to document and share computational workfows, including for research publications. The reproducibility of computational aspects of research is a key component of scientifc reproducibility but has not yet been assessed at scale for Jupyter notebooks associated with biomedical publications. Approach: We address computational reproducibility at 2 levels: (i) using fully automated workfows, we analyzed the computational reproducibility of Jupyter notebooks associated with publications indexed in the biomedical literature repository PubMed Central. We identifed such notebooks by mining the article's full text, trying to locate them on GitHub, and attempting to rerun them in an environment as close to the original as possible. We documented reproduction success and exceptions and explored relationships between notebook reproducibility and variables related to the notebooks or publications. (ii) This study represents a reproducibility attempt in and of itself, using essentially the same methodology twice on PubMed Central over the course of 2 years, during which the corpus of Jupyter notebooks from articles indexed in PubMed Central has grown in a highly dynamic fashion. Results: Out of 27,271 Jupyter notebooks from 2,660 GitHub repositories associated with 3,467 publications, 22,578 notebooks were written in Python, including 15,817 that had their dependencies declared in standard requirement fles and that we attempted to rerun automatically. For 10,388 of these, all declared dependencies could be installed successfully, and we reran them to assess reproducibility. Of these, 1,203 notebooks ran through without any errors, including 879 that produced results identical to those reported in the original notebook and 324 for which our results differed from the originally reported ones. Running the other notebooks resulted in exceptions. Conclusions: We zoom in on common problems and practices, highlight trends, and discuss potential improvements to Jupyter-related workfows associated with biomedical publications.

Author supplied keywords

Cite

CITATION STYLE

APA

Samuel, S., & Mietchen, D. (2024). Computational reproducibility of Jupyter notebooks from biomedical publications. GigaScience, 13. https://doi.org/10.1093/gigascience/giad113

Computational reproducibility of Jupyter notebooks from biomedical publications

Abstract

Author supplied keywords

Cite

Register to see more suggestions