GPTrans: A Biological Language Model-Based Approach for Predicting Disease-Associated Mutations in G Protein-Coupled Receptors

被引:0
|
作者
Wang, Xiaohua [1 ]
Zhang, Ming [1 ]
Yang, Xibei [1 ]
Yu, Dong-Jun [2 ]
Ge, Fang [3 ,4 ]
机构
[1] Jiangsu Univ Sci & Technol, Sch Comp, Zhenjiang 212100, Jiangsu, Peoples R China
[2] Nanjing Univ Sci & Technol, Sch Comp Sci & Engn, Nanjing 210094, Peoples R China
[3] Nanjing Univ Posts & Telecommun, State Key Lab Organ Elect & Informat Displays, Nanjing 210023, Peoples R China
[4] Nanjing Univ Posts & Telecommun, Inst Adv Mat IAM, Nanjing 210023, Peoples R China
基金
中国国家自然科学基金;
关键词
AMINO-ACID SUBSTITUTIONS; DRUG DISCOVERY; VARIANTS; SERVER;
D O I
10.1021/acs.jcim.4c01999
中图分类号
R914 [药物化学];
学科分类号
100701 ;
摘要
Accurately predicting mutations in G protein-coupled receptors (GPCRs) is critical for advancing disease diagnosis and drug discovery. In response to this imperative, GPTrans has emerged as a highly accurate predictor of disease-related mutations in GPCRs. The core innovation of GPTrans resides in the design of a novel feature extraction network, that is capable of integrating features from both wildtype and mutant protein variant sites, utilizing multifeature connections within a transformer framework to ensure comprehensive feature extraction. A key aspect of GPTrans's effectiveness is our introduction of an innovative deep feature integration strategy, which merges embeddings and class tokens from multiple protein language models, including evolutionary scale modeling and ProtTrans, thus shedding light on the biochemical properties of proteins. Leveraging transformer components and a self-attention mechanism, GPTrans captures higher-level representations of protein features. Employing both wildtype and mutation site information for feature fusion not only enriches the predictive feature set but also avoids the common issue of overestimation associated with sequence-based predictions. This approach distinguishes GPTrans, enabling it to significantly outperform existing methods. Our evaluations across diverse GPCR data sets, including ClinVar and MutHTP, demonstrate GPTrans's superior performance, with average AUC values of 0.874 and 0.590 in 10-fold cross-validation. Notably, compared to the AlphaMissense method, GPTrans exhibited a remarkable 38.03% improvement in accuracy when predicting disease-associated mutations in the MutHTP data set. A thorough analysis of the predicted results further validates the model's effectiveness. The source code, data sets, and prediction results for GPTrans are available for academic use at https://github.com/EduardWang/GPTrans.
引用
收藏
页码:9626 / 9642
页数:17
相关论文
共 50 条
  • [31] Interaction of membrane cholesterol with G Protein-Coupled Receptors: A multidimensional approach
    Chattopadhyay, A.
    EUROPEAN BIOPHYSICS JOURNAL WITH BIOPHYSICS LETTERS, 2013, 42 : S145 - S145
  • [32] A Hybrid Approach to Structure and Function Modeling of G Protein-Coupled Receptors
    Latek, Dorota
    Bajda, Marek
    Filipek, Slawomir
    JOURNAL OF CHEMICAL INFORMATION AND MODELING, 2016, 56 (04) : 630 - 641
  • [33] Targeting of G protein-coupled receptors to the plasma membrane in health and disease
    Ulloa-Aguirre, Alfredo
    Conn, P. Michael
    FRONTIERS IN BIOSCIENCE-LANDMARK, 2009, 14 : 973 - 994
  • [34] Adhesion G protein-coupled receptors in nervous system development and disease
    Tobias Langenhan
    Xianhua Piao
    Kelly R. Monk
    Nature Reviews Neuroscience, 2016, 17 : 550 - 561
  • [35] Targeting G Protein-Coupled Receptors in the Treatment of Parkinson's Disease
    Jones-Tabah, Jace
    JOURNAL OF MOLECULAR BIOLOGY, 2023, 435 (12)
  • [36] The role of G protein-coupled receptors in the pathology of Alzheimer's disease
    Thathiah, Amantha
    De Strooper, Bart
    NATURE REVIEWS NEUROSCIENCE, 2011, 12 (02) : 73 - 87
  • [37] The role of G protein-coupled receptors in the pathology of Alzheimer's disease
    Amantha Thathiah
    Bart De Strooper
    Nature Reviews Neuroscience, 2011, 12 : 73 - 87
  • [38] Adhesion G protein-coupled receptors in nervous system development and disease
    Langenhan, Tobias
    Piao, Xianhua
    Monk, Kelly R.
    NATURE REVIEWS NEUROSCIENCE, 2016, 17 (09) : 550 - 561
  • [39] Efficient Prediction of the Effect of Mutations on the Activation Kinetics of G Protein-Coupled Receptors Using a Maximum Caliber Approach
    Ramsey, Steven
    Provasi, Davide
    Moeller, Jan
    Lohse, Martin
    Filizola, Marta
    BIOPHYSICAL JOURNAL, 2020, 118 (03) : 92A - 93A
  • [40] G protein-coupled receptors (GPCRS)-directed mining: A pharmacophore-based approach.
    Nguyen, TB
    ABSTRACTS OF PAPERS OF THE AMERICAN CHEMICAL SOCIETY, 2003, 226 : U449 - U450