DocumentCode
2977571
Title
Preliminary Search Engine for Open Protein Identification
Author
Wenli Zhang ; Hao Chi ; Yuanzheng Lu ; Yuqing Huang ; Xiaofang Zhao ; Simin He
Author_Institution
Inst. of Comput. Technol., Beijing, China
fYear
2012
fDate
14-16 Dec. 2012
Firstpage
410
Lastpage
415
Abstract
Protein identification is the most important and basic problem for proteomics. Using tandem mass spectrometry and database search is one of the most widely used identification techniques. However, the improved sensitivity of mass spectrometers, rapid expansion of databases and more complex analysis, like post-translational modification and non-specific enzymatic digestion, have challenged current restricted protein identification search engines in scale and speed severely. In this paper, we proposed an open protein identification method relaxing enzyme, and presented our distributed design to support big protein database with non-specific digestion analysis based on pFind, a practical tandem mass spectra search engine developed in China. With classical bigger protein databases ipi. HUMAN and uniprot-sprot we got nearly linear speedup in a 20-blade cluster. By further analysis, we can expect real time identification to some extent.
Keywords
biology computing; database management systems; mass spectra; proteins; search engines; 20-blade cluster; big protein database support; database search; digestion analysis; open protein identification; preliminary search engine; proteomics; tandem mass spectrometry; Amino acids; Computers; Indexes; Peptides; Proteins; Search engines; cluster; distributed; non-specific enzymatic digestion; open protein identification; search engine;
fLanguage
English
Publisher
ieee
Conference_Titel
Parallel and Distributed Computing, Applications and Technologies (PDCAT), 2012 13th International Conference on
Conference_Location
Beijing
Print_ISBN
978-0-7695-4879-1
Type
conf
DOI
10.1109/PDCAT.2012.112
Filename
6589313
Link To Document