Abstract
Summary: A MapReduce-based implementation called MRMSPolygraph for parallelizing peptide identification from mass spectrometry data is presented. The underlying serial method, MSPolygraph, uses a novel hybrid approach to match an experimental spectrum against a combination of a protein sequence database and a spectral library. Our MapReduce implementation can run on any Hadoop cluster environment. Experimental results demonstrate that, relative to the serial version, MR-MSPolygraph reduces the time to solution from weeks to hours, for processing tens of thousands of experimental spectra. Speedup and other related performance studies are also reported on a 400-core Hadoop cluster using spectral datasets from environmental microbial communities as inputs. © The Author(s) 2011. Published by Oxford University Press.
Cite
CITATION STYLE
Kalyanaraman, A., Cannon, W. R., Latt, B., & Baxter, D. J. (2011). Mapreduce implementation of a hybrid spectral library-database search method for large-scale peptide identification. Bioinformatics, 27(21), 3072–3073. https://doi.org/10.1093/bioinformatics/btr523
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.