Identifying differences in protein expression levels by spectral counting and feature selection

被引:58
|
作者
Carvalho, P. C. [1 ]
Hewel, J. [2 ]
Barbosa, V. C. [1 ]
Yates, J. R., III [2 ]
机构
[1] Univ Fed Rio de Janeiro, COPPE, Programa Engn Sistemas & Computacao, BR-21945 Rio De Janeiro, Brazil
[2] Scripps Res Inst, Dept Cell Biol, La Jolla, CA USA
来源
GENETICS AND MOLECULAR RESEARCH | 2008年 / 7卷 / 02期
关键词
MudPIT; feature selection; support vector machine; spectral counting; feature ranking;
D O I
10.4238/vol7-2gmr426
中图分类号
Q5 [生物化学]; Q7 [分子生物学];
学科分类号
071010 ; 081704 ;
摘要
Spectral counting is a strategy to quantify relative protein concentrations in pre-digested protein mixtures analyzed by liquid chromatography online with tandem mass spectrometry. In the present study, we used combinations of normalization and statistical (feature selection) methods on spectral counting data to verify whether we could pinpoint which and how many proteins were differentially expressed when comparing complex protein mixtures. These combinations were evaluated on real, but controlled, experiments (yeast lysates were spiked with protein markers at different concentrations to simulate differences), which were therefore verifiable. The following normalization methods were applied: total signal, Z-normalization, hybrid normalization, and log preprocessing. The feature selection methods were: the Golub index, the Student t-test, a strategy based on the weighting used in a forward-support vector machine (SVM-F) model, and SVM recursive feature elimination. The results showed that Z-normalization combined with SVM-F correctly identified which and how many protein markers were added to the yeast lysates for all different concentrations. The software we used is available at http://pcarvalho.com/patternlab.
引用
收藏
页码:342 / 356
页数:15
相关论文
共 50 条
  • [21] Constraint Score Evaluation for Spectral Feature Selection
    Kalakech, Mariam
    Biela, Philippe
    Hamad, Denis
    Macaire, Ludovic
    NEURAL PROCESSING LETTERS, 2013, 38 (02) : 155 - 175
  • [22] Efficient Spectral Feature Selection with Minimum Redundancy
    Zhao, Zheng
    Wang, Lei
    Liu, Huan
    PROCEEDINGS OF THE TWENTY-FOURTH AAAI CONFERENCE ON ARTIFICIAL INTELLIGENCE (AAAI-10), 2010, : 673 - 678
  • [23] Feature Selection Using Counting Grids: Application to Microarray Data
    Lovato, Pietro
    Bicego, Manuele
    Cristani, Marco
    Jojic, Nebojsa
    Perina, Alessandro
    STRUCTURAL, SYNTACTIC, AND STATISTICAL PATTERN RECOGNITION, 2012, 7626 : 629 - 637
  • [24] Proteomic Analysis of Acetaminophen-Induced Changes in Mitochondrial Protein Expression Using Spectral Counting
    Stamper, Brendan D.
    Mohar, Isaac
    Kavanagh, Terrance J.
    Nelson, Sidney D.
    CHEMICAL RESEARCH IN TOXICOLOGY, 2011, 24 (04) : 549 - 558
  • [25] Unsupervised feature selection based on spectral regression from manifold learning for facial expression recognition
    Wang, Li
    Wang, Ke
    Li, Ruifeng
    IET COMPUTER VISION, 2015, 9 (05) : 655 - 662
  • [26] Unsupervised feature selection via discrete spectral clustering and feature weights
    Shang, Ronghua
    Kong, Jiarui
    Wang, Lujuan
    Zhang, Weitong
    Wang, Chao
    Li, Yangyang
    Jiao, Licheng
    NEUROCOMPUTING, 2023, 517 : 106 - 117
  • [27] A Feature Selection Technique based on Distributional Differences
    Kim, Sung-Dong
    JOURNAL OF INFORMATION PROCESSING SYSTEMS, 2006, 2 (01): : 23 - 27
  • [28] Feature extraction of protein expression levels based on classification of functional foods with SOM
    Fukushima, Tamon
    Yamamori, Kunihito
    Yoshihara, Ikuo
    Nagahama, Kiyoko
    ARTIFICIAL LIFE AND ROBOTICS, 2009, 13 (02) : 543 - 546
  • [29] Feature selection in multiword expression recognition
    Metin, Senem Kumova
    EXPERT SYSTEMS WITH APPLICATIONS, 2018, 92 : 106 - 123
  • [30] A Method of Identifying Airdrome Runways based on the Spectral and Structural Feature
    Guo Jianxing
    Liu Songlin
    Xu Honggen
    Liu Weimin
    ICCSE 2008: PROCEEDINGS OF THE THIRD INTERNATIONAL CONFERENCE ON COMPUTER SCIENCE & EDUCATION: ADVANCED COMPUTER TECHNOLOGY, NEW EDUCATION, 2008, : 1223 - 1226