Multivariate approach for protein identification based on mass spectrometric data

Citations

WEB OF SCIENCE

0
Citations

SCOPUS

0

초록

Protein mass spectrometry provides a powerful tool for detecting and identifying proteins. Several database searching algorithms may be used for this purpose. However, most of them depend on the heuristic approaches and the use of probability-based or statistical approach was very restrictive in the current algorithms. In this study, we present a statistical modelling of scores based on a generalized linear mixed model and provide a feasible computation method using penalized generalized weighted least squares. This model incorporates the dependency among matches into a new statistical scoring function, and uses the beta-binomial distribution to derive the score. Based on simulation experiments and analysis using real examples, we have improved protein searching performance and provided feasible computation procedures to deal with very large datasets. In particular, our methods may significantly increase accuracy in identifying medium and small proteins.

키워드

protein identification; mass spectrometry; generalized linear mixed model; two-part model; STRATEGIES; OLAV
제목
Multivariate approach for protein identification based on mass spectrometric data
저자
Lee, Jung Bok; Lee, Jae Won
DOI
10.1177/0962280211434960
발행일
2013-12
유형
Article
저널명
Statistical Methods in Medical Research
권
22
호
6
페이지
553 ~ 566