G4Boost: a machine learning-based tool for quadruplex identification and stability prediction

被引:14
作者
Cagirici, H. Busra [1 ]
Budak, Hikmet [2 ]
Sen, Taner Z. [1 ]
机构
[1] USDA ARS, Crop Improvement Genet Res Unit, Western Reg Res Ctr, 800 Buchanan St, Albany, CA 94710 USA
[2] Montana BioAgr Inc, Missoula, MT USA
关键词
G-quadruplex; Machine learning; Topology; Stability; Energy; Plants; Humans; RNA G-QUADRUPLEXES; SECONDARY STRUCTURE; WEB SERVER; DNA; PROMOTER; TRANSLATION; INHIBITION; PREVALENCE; TELOMERASE; SEQUENCE;
D O I
10.1186/s12859-022-04782-z
中图分类号
Q5 [生物化学];
学科分类号
071010 ; 081704 ;
摘要
Background G-quadruplexes (G4s), formed within guanine-rich nucleic acids, are secondary structures involved in important biological processes. Although every G4 motif has the potential to form a stable G4 structure, not every G4 motif would, and accurate energy-based methods are needed to assess their structural stability. Here, we present a decision tree-based prediction tool, G4Boost, to identify G4 motifs and predict their secondary structure folding probability and thermodynamic stability based on their sequences, nucleotide compositions, and estimated structural topologies. Results G4Boost predicted the quadruplex folding state with an accuracy greater then 93% and an F1-score of 0.96, and the folding energy with an RMSE of 4.28 and R-2 of 0.95 only by the means of sequence intrinsic feature. G4Boost was successfully applied and validated to predict the stability of experimentally-determined G4 structures, including for plants and humans. Conclusion G4Boost outperformed the three machine-learning based prediction tools, DeepG4, Quadron, and G4RNA Screener, in terms of both accuracy and F1-score, and can be highly useful for G4 prediction to understand gene regulation across species including plants and humans.
引用
收藏
页数:18
相关论文
共 50 条
[31]   Machine learning-based prediction of FeNi nanoparticle magnetization [J].
Williamson, Federico ;
Naciff, Nadhir ;
Catania, Carlos ;
dos Santos, Gonzalo ;
Amigo, Nicolas ;
Bringa, Eduardo M. .
JOURNAL OF MATERIALS RESEARCH AND TECHNOLOGY-JMR&T, 2024, 33 :5263-5276
[32]   Interpretability of machine learning-based prediction models in healthcare [J].
Stiglic, Gregor ;
Kocbek, Primoz ;
Fijacko, Nino ;
Zitnik, Marinka ;
Verbert, Katrien ;
Cilar, Leona .
WILEY INTERDISCIPLINARY REVIEWS-DATA MINING AND KNOWLEDGE DISCOVERY, 2020, 10 (05)
[33]   Machine Learning-Based Approach for Hardware Faults Prediction [J].
Khalil, Kasem ;
Eldash, Omar ;
Kumar, Ashok ;
Bayoumi, Magdy .
IEEE TRANSACTIONS ON CIRCUITS AND SYSTEMS I-REGULAR PAPERS, 2020, 67 (11) :3880-3892
[34]   Machine Learning-Based Prediction of the Martensite Start Temperature [J].
Wentzien, Marcel ;
Koch, Marcel ;
Friedrich, Thomas ;
Ingber, Jerome ;
Kempka, Henning ;
Schmalzried, Dirk ;
Kunert, Maik .
STEEL RESEARCH INTERNATIONAL, 2024, 95 (10)
[35]   Machine Learning-based RSSI Prediction in Factory Environments [J].
Webber, Julian ;
Suga, Norisato ;
Ano, Susumu ;
Jou, Yafei ;
Mehbodniya, Abolfazl ;
Higashimori, Toshihide ;
Yano, Kazuto ;
Suzuki, Yoshinori .
PROCEEDINGS OF 2019 25TH ASIA-PACIFIC CONFERENCE ON COMMUNICATIONS (APCC), 2019, :195-200
[36]   Machine learning-based approaches for disease gene prediction [J].
Duc-Hau Le .
BRIEFINGS IN FUNCTIONAL GENOMICS, 2020, 19 (5-6) :350-363
[37]   Machine Learning-based Seismic Prediction of Building Structures [J].
Liu, Shuai ;
Peng, Hailiang ;
Deng, Xiaolu .
PROCEEDINGS OF 2024 INTERNATIONAL CONFERENCE ON MACHINE INTELLIGENCE AND DIGITAL APPLICATIONS, MIDA2024, 2024, :256-261
[38]   Machine Learning-Based Prediction of Stroke in Emergency Departments [J].
Abedi, Vida ;
Misra, Debdipto ;
Chaudhary, Durgesh ;
Avula, Venkatesh ;
Schirmer, Clemens M. ;
Li, Jiang ;
Zand, Ramin .
THERAPEUTIC ADVANCES IN NEUROLOGICAL DISORDERS, 2024, 17
[39]   Machine learning-based model for prediction of concrete strength [J].
Aswal, Vivek Singh ;
Singh, B. K. ;
Maheshwari, Rohit .
MULTISCALE AND MULTIDISCIPLINARY MODELING EXPERIMENTS AND DESIGN, 2025, 8 (01)
[40]   A Machine Learning-Based Approach for Crop Price Prediction [J].
Gururaj, H. L. ;
Janhavi, V. ;
Lakshmi, H. ;
Soundarya, B. C. ;
Paramesha, K. ;
Ramesh, B. ;
Rajendra, A. B. .
JOURNAL OF CIRCUITS SYSTEMS AND COMPUTERS, 2024, 33 (03)