Literature DB >> 17952874

A new strategy to filter out false positive identifications of peptides in SEQUEST database search results.

Jiyang Zhang1, Jianqi Li, Hongwei Xie, Yunping Zhu, Fuchu He.   

Abstract

Based on the randomized database method and a linear discriminant function (LDF) model, a new strategy to filter out false positive matches in SEQUEST database search results is proposed. Given an experiment MS/MS dataset and a protein sequence database, a randomized database is constructed and merged with the original database. Then, all MS/MS spectra are searched against the combined database. For each expected false positive rate (FPR), LDFs are constructed for different charge states and used to filter out the false positive matches from the normal database. In order to investigate the error of FPR estimation, the new strategy was applied to a reference dataset. As a result, the estimated FPR was very close to the actual FPR. While applied to a human K562 cell line dataset, which is a complicated dataset from real sample, more matches could be confirmed than the traditional cutoff-based methods at the same estimated FPR. Also, though most of the results confirmed by the LDF model were consistent with those of PeptideProphet, the LDF model could still provide complementary information. These results indicate that the new method can reliably control the FPR of peptide identifications and is more sensitive than traditional cutoff-based methods.

Entities:  

Mesh:

Substances:

Year:  2007        PMID: 17952874     DOI: 10.1002/pmic.200600929

Source DB:  PubMed          Journal:  Proteomics        ISSN: 1615-9853            Impact factor:   3.984


  7 in total

1.  Bayesian nonparametric model for the validation of peptide identification in shotgun proteomics.

Authors:  Jiyang Zhang; Jie Ma; Lei Dou; Songfeng Wu; Xiaohong Qian; Hongwei Xie; Yunping Zhu; Fuchu He
Journal:  Mol Cell Proteomics       Date:  2008-11-12       Impact factor: 5.911

2.  Performance comparisons of nano-LC systems, electrospray sources and LC-MS-MS platforms.

Authors:  Qian Liu; Jennifer S Cobb; Joshua L Johnson; Qi Wang; Jeffrey N Agar
Journal:  J Chromatogr Sci       Date:  2013-01-17       Impact factor: 1.618

3.  Full-Featured, Real-Time Database Searching Platform Enables Fast and Accurate Multiplexed Quantitative Proteomics.

Authors:  Devin K Schweppe; Jimmy K Eng; Qing Yu; Derek Bailey; Ramin Rad; Jose Navarrete-Perea; Edward L Huttlin; Brian K Erickson; Joao A Paulo; Steven P Gygi
Journal:  J Proteome Res       Date:  2020-04-06       Impact factor: 4.466

4.  IDPicker 2.0: Improved protein assembly with high discrimination peptide identification filtering.

Authors:  Ze-Qiang Ma; Surendra Dasari; Matthew C Chambers; Michael D Litton; Scott M Sobecki; Lisa J Zimmerman; Patrick J Halvey; Birgit Schilling; Penelope M Drake; Bradford W Gibson; David L Tabb
Journal:  J Proteome Res       Date:  2009-08       Impact factor: 4.466

5.  Learning from decoys to improve the sensitivity and specificity of proteomics database search results.

Authors:  Amit Kumar Yadav; Dhirendra Kumar; Debasis Dash
Journal:  PLoS One       Date:  2012-11-26       Impact factor: 3.240

Review 6.  Bioinformatics in China: a personal perspective.

Authors:  Liping Wei; Jun Yu
Journal:  PLoS Comput Biol       Date:  2008-04-25       Impact factor: 4.475

7.  A nonparametric model for quality control of database search results in shotgun proteomics.

Authors:  Jiyang Zhang; Jianqi Li; Xin Liu; Hongwei Xie; Yunping Zhu; Fuchu He
Journal:  BMC Bioinformatics       Date:  2008-01-21       Impact factor: 3.169

  7 in total

北京卡尤迪生物科技股份有限公司 © 2022-2023.