Common data models to streamline metabolomics processing and annotation, and implementation in a Python pipeline

5Citations
Citations of this article
16Readers
Mendeley users who have this article in their library.
Get full text

Abstract

To standardize metabolomics data analysis and facilitate future computational developments, it is essential to have a set of well-defined templates for common data structures. Here we describe a collection of data structures involved in metabolomics data processing and illustrate how they are utilized in a full-featured Python-centric pipeline. We demonstrate the performance of the pipeline, and the details in annotation and quality control using largescale LC-MS metabolomics and lipidomics data and LC-MS/MS data. Multiple previously published datasets are also reanalyzed to showcase its utility in biological data analysis. This pipeline allows users to streamline data processing, quality control, annotation, and standardization in an efficient and transparent manner. This work fills a major gap in the Python ecosystem for computational metabolomics.

Cite

CITATION STYLE

APA

Mitchell, J. M., Chi, Y., Thapa, M., Pang, Z., Xia, J., & Li, S. (2024). Common data models to streamline metabolomics processing and annotation, and implementation in a Python pipeline. PLoS Computational Biology, 20(6 June). https://doi.org/10.1371/journal.pcbi.1011912

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free