2015Journal of Analytical ChemistryRequires access

Peptide identification in “shotgun” proteomics using tandem mass spectrometry: Comparison of search engine algorithms

Mark V. Ivanov, Lev I. Levitsky, Anna A. Lobas, Irina A. Tarasova, M. L. Pridatchenko, Victor G. Zgoda, Sergei A. Moshkovskii, Goran Mitulović, Mikhail V. Gorshkov

Open publisher page 2 citations

Abstract

High-throughput proteomics technologies are gaining popularity in different areas of life sciences. One of the main objectives of proteomics is characterization of the proteins in biological samples using liquid chromatography/mass spectrometry analysis of the corresponding proteolytic peptide mixtures. Both the complexity and the scale of experimental data obtained even from a single experimental run require specialized bioinformatic tools for automated data mining. One of the most important tools is a so-called proteomics search engine used for identification of proteins present in a sample by comparing experimental and theoretical tandem mass spectra. The latter are generated for the proteolytic peptides derived from a protein database. Peptide identifications obtained with the search engine are then scored according to the probability of a correct peptide-spectrum match. The purpose of this work was to perform a comparison of different search algorithms using data acquired for complex protein mixtures, including both annotated protein standards and clinical samples. The comparison was performed for three popular search engines: commercially available Mascot, as well as open-source X!Tandem and OMSSA. It was shown that the search engine OMSSA identifies in general a smaller number of proteins, while X!Tandem and Mascot deliver similar performance. We found no compelling reasons for using the commercial search engine instead of its open source competitor.

About this research paper

What this paper is about

High-throughput proteomics technologies are gaining popularity in different areas of life sciences. One of the main objectives of proteomics is characterization of the proteins in biological samples using liquid chromatography/mass spectrometry analysis of the corresponding proteolytic peptide mixtures. Both the complexity and the scale of experimental data obtained even from a single experimental run require specialized bioinformatic tools for automated data mining. One of the most important tools is a so-called proteomics search engine used for identification of proteins present in a sample by comparing experimental and theoretical tandem mass spectra. The latter are generated for the proteolytic peptides derived from a protein database. Peptide identifications obtained with the search engine are then scored according to the probability of a correct peptide-spectrum match. The purpose of this work was to perform a comparison of different search algorithms using data acquired for complex protein mixtures, including both annotated protein standards and clinical samples. The comparison was performed for three popular search engines: commercially available Mascot, as well as open-source X!Tandem and OMSSA. It was shown that the search engine OMSSA identifies in general a smaller number of proteins, while X!Tandem and Mascot deliver similar performance. We found no compelling reasons for using the commercial search engine instead of its open source competitor.

Why it matters

OpenAlex reports 2 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

High-throughput proteomics technologies are gaining popularity in different areas of life sciences. One of the main objectives of proteomics is characterization of the proteins in biological samples using liquid chromatography/mass spectrometry analysis of the corresponding proteolytic peptide mixtures. Both the complexity and the scale of experimental data obtained even from a single experimental run require specialized bioinformatic tools for automated data mining. One of the most important tools is a so-called proteomics search engine used for identification of proteins present in a sample by comparing experimental and theoretical tandem mass spectra. The latter are generated for the proteolytic peptides derived from a protein database. Peptide identifications obtained with the search engine are then scored according to the probability of a correct peptide-spectrum match. The purpose of this work was to perform a comparison of different search algorithms using data acquired for complex protein mixtures, including both annotated protein standards and clinical samples. The comparison was performed for three popular search engines: commercially available Mascot, as well as open-source X!Tandem and OMSSA. It was shown that the search engine OMSSA identifies in general a smaller number of proteins, while X!Tandem and Mascot deliver similar performance. We found no compelling reasons for using the commercial search engine instead of its open source competitor.

Key concepts: Mascot, Database search engine, Shotgun proteomics, Tandem mass spectrometry, Chemistry, Proteomics, Search engine, Mass spectrometry

Related papers

Back to paper searchBrowse research topicsOriginal source
Peptide identification in “shotgun” proteomics using tandem mass spectrometry: Comparison of search engine algorithms — Research Paper | ScholarLens