De Novo Sequencing and Homology Searching

peer-reviewed · Molecular & Cellular Proteomics · 2012

peer-reviewed · Molecular & Cellular Proteomics · 2012. Bin Ma et al. In proteomics, de novo sequencing is the process of deriving peptide sequences from tandem mass spectra…
Date 2012-02-01
Type peer-reviewed
Venue Molecular & Cellular Proteomics
Publisher Elsevier BV
Contribution review
DOI 10.1074/mcp.O111.014902
Citations (OpenAlex) 188
Venue 2-year citedness 4.17

Abstract

In proteomics, de novo sequencing is the process of deriving peptide sequences from tandem mass spectra without the assistance of a sequence database. Such analyses have traditionally been performed manually by human experts, and more recently by computer programs that have been developed because of the need for higher throughput. Although powerful, de novo sequencing often can only determine partially correct sequence tags because of imperfect tandem mass spectra. However, these sequence tags can then be searched in a sequence database to identify the exact or a homologous peptide. Homology searches are particularly useful for the study of organisms whose genomes have not been sequenced. This tutorial will present background important to understanding de novo sequencing, suggestions on how to do this manually, plus descriptions of computer algorithms used to automate this process and to subsequently carryout homology-based database searches. This Tutorial is part of the International Proteomics Tutorial Programme (IPTP 1).

Authors

  1. Bin Ma · Rapid Novor Inc., University of Waterloo, University of Western Ontario
  2. Richard S. Johnson · Immunex Corporation, Institute for Systems Biology, Massachusetts Institute of Technology

Methods and tools

Cites (25)

Cited by (20)

Seen in the charts

Back to the full map

Back to top