DocumentCode
2377781
Title
FISH: Finding of identical spectra set for Homogenous peptide using two-stage clustering algorithm
Author
Lee, Seungmook ; Kwon, Min-Seok ; Lee, Hyoung-Joo ; Paik, Young-Ki ; Lee, Jae K. ; Park, Taesung
Author_Institution
Dept. of Stat., Seoul Nat. Univ., Seoul, South Korea
fYear
2010
fDate
18-18 Dec. 2010
Firstpage
73
Lastpage
76
Abstract
MS/MS experiments generate hundreds to tens of thousands of fragment ion spectra during the experiment. In the peptide identification, MS/MS spectra are often identified by database searching algorithms such as SEQUEST and Mascot. Most database searching algorithms calculate score functions to compare the experimental MS/MS spectra with theoretical MS/MS spectra of certain peptides derived from protein sequence databases. However, these indirect methods are vulnerable to potential errors. Thus, some spectra may not be assigned to peptides. To overcome these limitations, we propose a novel algorithm called Finding Identical Spectra for Homogenous peptide (FISH) by using two-stage clustering algorithm. Our proposed FISH algorithm can cluster spectra from the same peptide as a group based on the direct comparison. That is, through all possible pair-wise comparisons, our FISH method provides a set of spectra from the same peptide. To investigate the efficiency of our proposed method, we performed Nano-LC-MS/MS experiment for human tissue samples. Also, we conducted the simulation study to compare the performance of our proposed method with other database searching methods. Our simulation study showed that FISH yielded higher sensitivity than the other methods.
Keywords
bioinformatics; mass spectra; organic compounds; pattern clustering; proteomics; query processing; search engines; FISH algorithm; Finding of Identical Spectra Set for Homogenous peptide; MS/MS spectra; Mascot; SEQUEST; database searching algorithms; fragment ion spectra; human tissue samples; peptide identification; two-stage clustering algorithm; component; direct comparison; hierarchical clustering; moving average; peptide identification;
fLanguage
English
Publisher
ieee
Conference_Titel
Bioinformatics and Biomedicine Workshops (BIBMW), 2010 IEEE International Conference on
Conference_Location
Hong, Kong
Print_ISBN
978-1-4244-8303-7
Electronic_ISBN
978-1-4244-8304-4
Type
conf
DOI
10.1109/BIBMW.2010.5703776
Filename
5703776
Link To Document