Harnessing pre-trained models for accurate prediction of protein-ligand binding affinity

被引:0
|
作者
Li, Jiashan [1 ]
Gong, Xinqi [1 ]
机构
[1] Renmin Univ China, Inst Math Sci, Sch Math, 59 Zhongguancun St, Beijing 100872, Peoples R China
来源
BMC BIOINFORMATICS | 2025年 / 26卷 / 01期
关键词
Binding affinity; Binding site prediction; Molecular representation; Molecular pre-training; SCORING FUNCTIONS; DOCKING; GLIDE;
D O I
10.1186/s12859-025-06064-w
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
BackgroundThe binding between proteins and ligands plays a crucial role in the field of drug discovery. However, this area currently faces numerous challenges. On one hand, existing methods are constrained by the limited availability of labeled data, often performing inadequately when addressing complex protein-ligand interactions. On the other hand, many models struggle to effectively capture the flexible variations and relative spatial relationships between proteins and ligands. These issues not only significantly hinder the advancement of protein-ligand binding research but also adversely affect the accuracy and efficiency of drug discovery. Therefore, in response to these challenges, our study aims to enhance predictive capabilities through innovative approaches, providing more reliable support for drug discovery efforts.MethodsThis study leverages a pre-trained model with spatial awareness to enhance the prediction of protein-ligand binding affinity. By perturbing the structures of small molecules in a manner consistent with physical constraints and employing self-supervised tasks, we improve the representation of small molecule structures, allowing for better adaptation to affinity predictions. Meanwhile, our approach enables the identification of potential binding sites on proteins.ResultsOur model demonstrates a significantly higher correlation coefficient in binding affinity predictions. Extensive evaluation on the PDBBind v2019 refined set, CASF, and Merck FEP benchmarks confirms the model's robustness and strong generalization across diverse datasets. Additionally, the model achieves over 95% in classification ROC for binding site identification, underscoring its high accuracy in pinpointing protein-ligand interaction regions.ConclusionThis research presents a novel approach that not only enhances the accuracy of binding affinity predictions but also facilitates the identification of binding sites, showcasing the potential of pre-trained models in computational drug design. Data and code are available at https://github.com/MIALAB-RUC/SableBind.
引用
收藏
页数:21
相关论文
共 50 条
  • [1] Protein-ligand binding affinity prediction model based on graph attention network
    Yuan, Hong
    Huang, Jing
    Li, Jin
    MATHEMATICAL BIOSCIENCES AND ENGINEERING, 2021, 18 (06) : 9148 - 9162
  • [2] Ensembling methods for protein-ligand binding affinity prediction
    Cader, Jiffriya Mohamed Abdul
    Newton, M. A. Hakim
    Rahman, Julia
    Cader, Akmal Jahan Mohamed Abdul
    Sattar, Abdul
    SCIENTIFIC REPORTS, 2024, 14 (01):
  • [3] DeepAtom: A Framework for Protein-Ligand Binding Affinity Prediction
    Li, Yanjun
    Rezaei, Mohammad A.
    Li, Chenglong
    Li, Xiaolin
    2019 IEEE INTERNATIONAL CONFERENCE ON BIOINFORMATICS AND BIOMEDICINE (BIBM), 2019, : 303 - 310
  • [4] Ensemble of local and global information for Protein-Ligand Binding Affinity Prediction
    Li, Gaili
    Yuan, Yongna
    Zhang, Ruisheng
    COMPUTATIONAL BIOLOGY AND CHEMISTRY, 2023, 107
  • [5] Importance of Ligand Reorganization Free Energy in Protein-Ligand Binding-Affinity Prediction
    Yang, Chao-Yie
    Sun, Haiying
    Chen, Jianyong
    Nikolovska-Coleska, Zaneta
    Wang, Shaomeng
    JOURNAL OF THE AMERICAN CHEMICAL SOCIETY, 2009, 131 (38) : 13709 - 13721
  • [6] GAABind: a geometry-aware attention-based network for accurate protein-ligand binding pose and binding affinity prediction
    Tan, Huishuang
    Wang, Zhixin
    Hu, Guang
    BRIEFINGS IN BIOINFORMATICS, 2024, 25 (01)
  • [7] Learning protein-ligand binding affinity with atomic environment vectors
    Meli, Rocco
    Anighoro, Andrew
    Bodkin, Mike J.
    Morris, Garrett M.
    Biggin, Philip C.
    JOURNAL OF CHEMINFORMATICS, 2021, 13 (01)
  • [8] Structure-based protein-ligand interaction fingerprints for binding affinity prediction
    Wang, Debby D.
    Chan, Moon-Tong
    Yan, Hong
    COMPUTATIONAL AND STRUCTURAL BIOTECHNOLOGY JOURNAL, 2021, 19 : 6291 - 6300
  • [9] Multi-task bioassay pre-training for protein-ligand binding affinity prediction
    Yan, Jiaxian
    Ye, Zhaofeng
    Yang, Ziyi
    Lu, Chengqiang
    Zhang, Shengyu
    Liu, Qi
    Qiu, Jiezhong
    BRIEFINGS IN BIOINFORMATICS, 2024, 25 (01)
  • [10] Structure-based, deep-learning models for protein-ligand binding affinity prediction
    Debby D. Wang
    Wenhui Wu
    Ran Wang
    Journal of Cheminformatics, 16