Prediction of peptide fragment ion mass spectra by data mining techniques.

Nai-ping Dong,Yi-Zeng Liang,Qing-song Xu,Daniel K W Mok,Lun-zhao Yi,Hong-mei Lu,Min He,Wei Fan
DOI: https://doi.org/10.1021/ac501094m
IF: 7.4
2014-01-01
Analytical Chemistry
Abstract:Accurate prediction of peptide fragment ion mass spectra is one of the critical factors to guarantee confident peptide identification by protein sequence database search in bottom-up proteomics. In an attempt to accurately and comprehensively predict this type of mass spectra, a framework named (MSPBPI)-P-2 is proposed. (MSPBPI)-P-2 first extracts fragment ions from large-scale MS/MS spectra data sets according to the peptide fragmentation pathways and uses binary trees to divide the obtained bulky data into tens to more than 1000 regions. For each adequate region, stochastic gradient boosting tree regression model is constructed. By constructing hundreds of these models, (MSPBPI)-P-2 is able to predict MS/MS spectra for unmodified and modified peptides with reasonable accuracy. Moreover, high consistency between predicted and experimental MS/MS spectra derived from different ion trap instruments with low and high resolving power is achieved. MS2PBPI outperforms existing algorithms MassAnalyzer and PeptideART.
What problem does this paper attempt to address?