Mapreduce implementation of a hybrid spectral library-database search method for large-scale peptide identification

24Citations
Citations of this article
47Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

Summary: A MapReduce-based implementation called MRMSPolygraph for parallelizing peptide identification from mass spectrometry data is presented. The underlying serial method, MSPolygraph, uses a novel hybrid approach to match an experimental spectrum against a combination of a protein sequence database and a spectral library. Our MapReduce implementation can run on any Hadoop cluster environment. Experimental results demonstrate that, relative to the serial version, MR-MSPolygraph reduces the time to solution from weeks to hours, for processing tens of thousands of experimental spectra. Speedup and other related performance studies are also reported on a 400-core Hadoop cluster using spectral datasets from environmental microbial communities as inputs. © The Author(s) 2011. Published by Oxford University Press.

Cite

CITATION STYLE

APA

Kalyanaraman, A., Cannon, W. R., Latt, B., & Baxter, D. J. (2011). Mapreduce implementation of a hybrid spectral library-database search method for large-scale peptide identification. Bioinformatics, 27(21), 3072–3073. https://doi.org/10.1093/bioinformatics/btr523

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free