Database search tools identify peptides by matching tandem mass spectra against a protein database. We study an alternative approach when all plausible de novo interpretations of a spectrum (spectral dictionary) are generated and then quickly matched against the database. We present a new MS-Dictionary algorithm for efficiently generating spectral dictionaries and demonstrate that MS-Dictionary can identify spectra that are missed in the database search. We argue that MS-Dictionary enables proteogenomics searches in six-frame translation of genomic sequences that may be prohibitively time-consuming for existing database search approaches. We show that such searches allow one to correct sequencing errors and find programmed frameshifts.
Spectral Dictionaries
Sangtae Kim,Nitin Gupta,N. Bandeira,P. Pevzner
Published 2009 in Molecular & Cellular Proteomics
ABSTRACT
PUBLICATION RECORD
- Publication year
2009
- Venue
Molecular & Cellular Proteomics
- Publication date
2009-01-01
- Fields of study
Biology, Medicine, Chemistry, Computer Science
- Identifiers
- External record
- Source metadata
Semantic Scholar, PubMed
CITATION MAP
EXTRACTION MAP
CLAIMS
- No claims are published for this paper.
CONCEPTS
- No concepts are published for this paper.
REFERENCES
Showing 1-57 of 57 references · Page 1 of 1
CITED BY
Showing 1-87 of 87 citing papers · Page 1 of 1