False Discovery Rates of Protein Identifications: A Strike against the Two-Peptide Rule
- 23 July 2009
- journal article
- research article
- Published by American Chemical Society (ACS) in Journal of Proteome Research
- Vol. 8 (9) , 4173-4181
- https://doi.org/10.1021/pr9004794
Abstract
Most proteomics studies attempt to maximize the number of peptide identifications and subsequently infer proteins containing two or more peptides as reliable protein identifications. In this study, we evaluate the effect of this “two-peptide” rule on protein identifications, using multiple search tools and data sets. Contrary to the intuition, the “two-peptide” rule reduces the number of protein identifications in the target database more significantly than in the decoy database and results in increased false discovery rates, compared to the case when single-hit proteins are not discarded. We therefore recommend that the “two-peptide” rule should be abandoned, and instead, protein identifications should be subject to the estimation of error rates, as is the case with peptide identifications. We further extend the generating function approach (originally proposed for evaluating matches between a peptide and a single spectrum) to evaluating matches between a protein and an entire spectral data set.Keywords
This publication has 29 references indexed in Scilit:
- A Bayesian Approach to Protein Inference Problem in Shotgun ProteomicsJournal of Computational Biology, 2009
- Spectral DictionariesMolecular & Cellular Proteomics, 2009
- Spectral Probabilities and Generating Functions of Tandem Mass Spectra: A Strike against Decoy DatabasesJournal of Proteome Research, 2008
- Comparative proteogenomics: Combining mass spectrometry and comparative genomics to analyze multiple genomesGenome Research, 2008
- Assigning Significance to Peptides Identified by Tandem Mass Spectrometry Using Decoy DatabasesJournal of Proteome Research, 2007
- Clustering Millions of Tandem Mass SpectraJournal of Proteome Research, 2007
- Whole proteome analysis of post-translational modifications: Applications of mass-spectrometry for proteogenomic annotationGenome Research, 2007
- Target-decoy search strategy for increased confidence in large-scale protein identifications by mass spectrometryNature Methods, 2007
- Mass spectrometry-based proteomicsNature, 2003
- Empirical Statistical Model To Estimate the Accuracy of Peptide Identifications Made by MS/MS and Database SearchAnalytical Chemistry, 2002