StackedEnC-AOP: prediction of antioxidant proteins using transform evolutionary and sequential features based multi-scale vector with stacked ensemble learning.
Antioxidant proteins
Evolutionary features
Feature selection
Prediction
Stacked ensemble model
Transformation
Journal
BMC bioinformatics
ISSN: 1471-2105
Titre abrégé: BMC Bioinformatics
Pays: England
ID NLM: 100965194
Informations de publication
Date de publication:
04 Aug 2024
04 Aug 2024
Historique:
received:
19
04
2024
accepted:
29
07
2024
medline:
5
8
2024
pubmed:
5
8
2024
entrez:
4
8
2024
Statut:
epublish
Résumé
Antioxidant proteins are involved in several biological processes and can protect DNA and cells from the damage of free radicals. These proteins regulate the body's oxidative stress and perform a significant role in many antioxidant-based drugs. The current invitro-based medications are costly, time-consuming, and unable to efficiently screen and identify the targeted motif of antioxidant proteins. In this model, we proposed an accurate prediction method to discriminate antioxidant proteins namely StackedEnC-AOP. The training sequences are formulation encoded via incorporating a discrete wavelet transform (DWT) into the evolutionary matrix to decompose the PSSM-based images via two levels of DWT to form a Pseudo position-specific scoring matrix (PsePSSM-DWT) based embedded vector. Additionally, the Evolutionary difference formula and composite physiochemical properties methods are also employed to collect the structural and sequential descriptors. Then the combined vector of sequential features, evolutionary descriptors, and physiochemical properties is produced to cover the flaws of individual encoding schemes. To reduce the computational cost of the combined features vector, the optimal features are chosen using Minimum redundancy and maximum relevance (mRMR). The optimal feature vector is trained using a stacking-based ensemble meta-model. Our developed StackedEnC-AOP method reported a prediction accuracy of 98.40% and an AUC of 0.99 via training sequences. To evaluate model validation, the StackedEnC-AOP training model using an independent set achieved an accuracy of 96.92% and an AUC of 0.98. Our proposed StackedEnC-AOP strategy performed significantly better than current computational models with a ~ 5% and ~ 3% improved accuracy via training and independent sets, respectively. The efficacy and consistency of our proposed StackedEnC-AOP make it a valuable tool for data scientists and can execute a key role in research academia and drug design.
Sections du résumé
BACKGROUND
BACKGROUND
Antioxidant proteins are involved in several biological processes and can protect DNA and cells from the damage of free radicals. These proteins regulate the body's oxidative stress and perform a significant role in many antioxidant-based drugs. The current invitro-based medications are costly, time-consuming, and unable to efficiently screen and identify the targeted motif of antioxidant proteins.
METHODS
METHODS
In this model, we proposed an accurate prediction method to discriminate antioxidant proteins namely StackedEnC-AOP. The training sequences are formulation encoded via incorporating a discrete wavelet transform (DWT) into the evolutionary matrix to decompose the PSSM-based images via two levels of DWT to form a Pseudo position-specific scoring matrix (PsePSSM-DWT) based embedded vector. Additionally, the Evolutionary difference formula and composite physiochemical properties methods are also employed to collect the structural and sequential descriptors. Then the combined vector of sequential features, evolutionary descriptors, and physiochemical properties is produced to cover the flaws of individual encoding schemes. To reduce the computational cost of the combined features vector, the optimal features are chosen using Minimum redundancy and maximum relevance (mRMR). The optimal feature vector is trained using a stacking-based ensemble meta-model.
RESULTS
RESULTS
Our developed StackedEnC-AOP method reported a prediction accuracy of 98.40% and an AUC of 0.99 via training sequences. To evaluate model validation, the StackedEnC-AOP training model using an independent set achieved an accuracy of 96.92% and an AUC of 0.98.
CONCLUSION
CONCLUSIONS
Our proposed StackedEnC-AOP strategy performed significantly better than current computational models with a ~ 5% and ~ 3% improved accuracy via training and independent sets, respectively. The efficacy and consistency of our proposed StackedEnC-AOP make it a valuable tool for data scientists and can execute a key role in research academia and drug design.
Identifiants
pubmed: 39098908
doi: 10.1186/s12859-024-05884-6
pii: 10.1186/s12859-024-05884-6
doi:
Substances chimiques
Antioxidants
0
Proteins
0
Types de publication
Journal Article
Langues
eng
Sous-ensembles de citation
IM
Pagination
256Subventions
Organisme : National Natural Science Foundation of China
ID : 62131004
Organisme : National Key Research and Development Program of China
ID : 2022ZD0117700
Informations de copyright
© 2024. The Author(s).
Références
Phaniendra A, Jestadi DB, Periyasamy L. Free radicals: properties, sources, targets, and their implication in various diseases. Indian J Clin Biochem. 2015;30:11–26.
pubmed: 25646037
doi: 10.1007/s12291-014-0446-0
Sundaram Sanjay S, Shukla AK. Free radicals versus antioxidants. In: Sanjay SS, Shukla AK, editors. Potential therapeutic applications of nano-antioxidants. Springer: Singapore; 2021. p. 1–17.
doi: 10.1007/978-981-16-1143-8
Nimse SB, Pal D. Free radicals, natural antioxidants, and their reaction mechanisms. RSC Adv. 2015;5(35):27986–8006.
doi: 10.1039/C4RA13315C
Rajendran P, Nandakumar N, Rengarajan T, Palaniswami R, Gnanadhas EN, Lakshminarasaiah U, Gopas J, Nishigaki I. Antioxidants and human diseases. Clin Chim Acta. 2014;436:332–47.
pubmed: 24933428
doi: 10.1016/j.cca.2014.06.004
Jomova K, Raptova R, Alomar SY, Alwasel SH, Nepovimova E, Kuca K, Valko M. Reactive oxygen species, toxicity, oxidative stress, and antioxidants: chronic diseases and aging. Arch Toxicol. 2023;97(10):2499–574.
pubmed: 37597078
pmcid: 10475008
doi: 10.1007/s00204-023-03562-9
Kıran TR, Otlu O, Karabulut AB. Oxidative stress and antioxidants in health and disease. J Lab Med. 2023;47(1):1–11.
doi: 10.1515/labmed-2022-0108
He P, Zhang Y, Zhang Y, Zhang L, Lin Z, Sun C, Wu H, Zhang M. Isolation, identification of antioxidant peptides from earthworm proteins and analysis of the structure–activity relationship of the peptides based on quantum chemical calculations. Food Chem. 2024;431:137137.
pubmed: 37591140
doi: 10.1016/j.foodchem.2023.137137
Pagan LU, Gomes MJ, Gatto M, Mota GA, Okoshi K, Okoshi MP. The role of oxidative stress in the aging heart. Antioxidants. 2022;11(2):336.
pubmed: 35204217
pmcid: 8868312
doi: 10.3390/antiox11020336
Chang K-H, Chen C-M. The role of oxidative stress in Parkinson’s disease. Antioxidants. 2020;9(7):597.
pubmed: 32650609
pmcid: 7402083
doi: 10.3390/antiox9070597
Sun Q, Kong W, Mou X, Wang S. Transcriptional regulation analysis of Alzheimer’s disease based on FastNCA algorithm. Curr Bioinform. 2019;14(8):771–82.
doi: 10.2174/1574893614666190919150411
Liguori I, Russo G, Curcio F, Bulli G, Aran L, Della-Morte D, Gargiulo G, Testa G, Cacciatore F, Bonaduce D. Oxidative stress, aging, and diseases. Clin Interv Aging. 2018;13:757–72.
pubmed: 29731617
pmcid: 5927356
doi: 10.2147/CIA.S158513
Reddy VP. Oxidative stress in health and disease. Biomedicines. 2023;11(11):2925.
pubmed: 38001926
pmcid: 10669448
doi: 10.3390/biomedicines11112925
Li X, Tang Q, Tang H, Chen W. Identifying antioxidant proteins by combining multiple methods. Front Bioeng Biotechnol. 2020;8:858.
pubmed: 32793581
pmcid: 7391787
doi: 10.3389/fbioe.2020.00858
Chaudhary P, Janmeda P, Docea AO, Yeskaliyeva B, Abdull Razis AF, Modu B, Calina D, Sharifi-Rad J. Oxidative stress, free radicals and antioxidants: potential crosstalk in the pathophysiology of human diseases. Front Chem. 2023;11:1158198.
pubmed: 37234200
pmcid: 10206224
doi: 10.3389/fchem.2023.1158198
Dhalaria R, Verma R, Kumar D, Puri S, Tapwal A, Kumar V, Nepovimova E, Kuca K. Bioactive compounds of edible fruits with their anti-aging properties: a comprehensive review to prolong human life. Antioxidants. 2020;9(11):1123.
pubmed: 33202871
pmcid: 7698232
doi: 10.3390/antiox9111123
Moulahoum H, Ghorbanizamani F, Timur S, Zihnioglu F. Beyond natural antioxidants in cancer therapy: novel synthetic approaches in harnessing oxidative stress. In: Chakraborti S, editor. Handbook of oxidative stress in cancer: therapeutic aspects. Springer: Singapore; 2022. p. 1–17.
Rojas-Fernandez CH, Tyber K. Benefits, potential harms, and optimal use of nutritional supplementation for preventing progression of age-related macular degeneration. Ann Pharmacother. 2017;51(3):264–70.
pubmed: 27866147
doi: 10.1177/1060028016680643
Mishra N, Tripathi S, Nahar L, Sarker SD, Kumar A: Mitigation of arsenic poisoning induced oxidative stress and genotoxicity by Ocimum gratissimum L. Toxicon 2024:107603.
Pisoschi AM, Negulescu GP. Methods for total antioxidant activity determination: a review. Biochem Anal Biochem. 2011;1(1):106.
Wachirattanapongmetee K, Katekaew S, Weerapreeyakul N, Thawornchinsombut S. Differentiation of protein types extracted from tilapia byproducts by FTIR spectroscopy combined with chemometric analysis and their antioxidant protein hydrolysates. Food Chem. 2024;437:137862.
pubmed: 37931446
doi: 10.1016/j.foodchem.2023.137862
Madhani Mohammed Sadhakathullah AH, Paulo Mirasol S, Molina García BG, Torras Costa J, Armelín Diggroc EA. PLA-PEG-cholesterol biomimetic membrane for electrochemical sensing of antioxidants. Electrochim Acta. 2024;476:143716.
doi: 10.1016/j.electacta.2023.143716
Chen L, Chen S, Rong Y, Zeng W, Hu Z, Ma X, Feng S. Identification and evaluation of antioxidant peptides from highland barley distiller’s grains protein hydrolysate assisted by molecular docking. Food Chem. 2024;434:137441.
pubmed: 37769603
doi: 10.1016/j.foodchem.2023.137441
Li W, Zhu L, Zhang F, Han C, Li P, Jiang J. A novel strategy by combining foam fractionation with high-speed countercurrent chromatography for the rapid and efficient isolation of antioxidants and cytostatics from Camellia oleifera cake. Food Res Int. 2024;176:113798.
pubmed: 38163709
doi: 10.1016/j.foodres.2023.113798
Lv Z, Cui F, Zou Q, Zhang L, Xu L. Anticancer peptides prediction with deep representation learning features. Brief Bioinform. 2021;22(5):bbab008.
pubmed: 33529337
doi: 10.1093/bib/bbab008
Lv Z, Zhang J, Ding H, Zou Q. RF-PseU: a random forest predictor for RNA pseudouridine sites. Front Bioeng Biotechnol. 2020;8:134.
pubmed: 32175316
pmcid: 7054385
doi: 10.3389/fbioe.2020.00134
Lv H, Dao F-Y, Zulfiqar H, Lin H. DeepIPs: comprehensive assessment and computational identification of phosphorylation sites of SARS-CoV-2 infection using a deep learning-based approach. Brief Bioinform. 2021;22(6):bbab244.
pubmed: 34184738
doi: 10.1093/bib/bbab244
Olawoye B, Fagbohun OF, Popoola-Akinola O, Akinsola JET, Akanbi CT. A supervised machine learning approach for the prediction of antioxidant activities of Amaranthus viridis seed. Heliyon. 2024;10:e24506.
pubmed: 38322916
pmcid: 10844001
doi: 10.1016/j.heliyon.2024.e24506
Meng C, Pei Y, Bu Y, Zou Q, Ju Y. Machine learning-based antioxidant protein identification model: progress and evaluation. J Cell Biochem. 2023;124:1825–34.
pubmed: 37877550
doi: 10.1002/jcb.30491
Feng P, Ding H, Lin H, Chen W. AOD: the antioxidant protein database. Sci Rep. 2017;7(1):7449.
pubmed: 28784999
pmcid: 5547145
doi: 10.1038/s41598-017-08115-6
Fernández-Blanco E, Aguiar-Pulido V, Munteanu CR, Dorado J. Random Forest classification based on star graph topological indices for antioxidant proteins. J Theor Biol. 2013;317:331–7.
pubmed: 23116665
doi: 10.1016/j.jtbi.2012.10.006
Feng P-M, Lin H, Chen W. Identification of antioxidants from sequence information using naive Bayes. Comput Math Methods Med. 2013;2013:567529.
pubmed: 24062796
pmcid: 3766563
doi: 10.1155/2013/567529
Feng P, Chen W, Lin H. Identifying antioxidant proteins by using optimal dipeptide compositions. Interdiscip Sci Comput Life Sci. 2016;8:186–91.
doi: 10.1007/s12539-015-0124-9
Zhang L, Zhang C, Gao R, Yang R. Incorporating g-gap dipeptide composition and position specific scoring matrix for identifying antioxidant proteins. In: 2015 IEEE 28th Canadian conference on electrical and computer engineering (CCECE). IEEE; 2015. p. 31–6.
Zhang L, Zhang C, Gao R, Yang R, Song Q. Sequence based prediction of antioxidant proteins using a classifier selection strategy. PLoS ONE. 2016;11(9):e0163274.
pubmed: 27662651
pmcid: 5035026
doi: 10.1371/journal.pone.0163274
Xu L, Liang G, Shi S, Liao C. SeqSVM: a sequence-based support vector machine method for identifying antioxidant proteins. Int J Mol Sci. 2018;19(6):1773.
pubmed: 29914044
pmcid: 6032279
doi: 10.3390/ijms19061773
Meng C, Jin S, Wang L, Guo F, Zou Q. AOPs-SVM: a sequence-based classifier of antioxidant proteins using a support vector machine. Front Bioeng Biotechnol. 2019;7:224.
pubmed: 31620433
pmcid: 6759716
doi: 10.3389/fbioe.2019.00224
Butt AH, Rasool N, Khan YD. Prediction of antioxidant proteins by incorporating statistical moments based features into Chou’s PseAAC. J Theor Biol. 2019;473:1–8.
pubmed: 31005614
doi: 10.1016/j.jtbi.2019.04.019
Ahmad A, Akbar S, Hayat M, Ali F, Khan S, Sohail M. Identification of antioxidant proteins using a discriminative intelligent model of k-space amino acid pairs based descriptors incorporating with ensemble feature selection. Biocybern Biomed Eng. 2022;42(2):727–35.
doi: 10.1016/j.bbe.2020.10.003
Ho Thanh Lam L, Le NH, Van Tuan L, Tran Ban H, Nguyen Khanh Hung T, Nguyen NTK, Huu Dang L, Le NQK. Machine learning model for identifying antioxidant proteins using features calculated from primary sequences. Biology. 2020;9(10):325.
pubmed: 33036150
pmcid: 7599600
doi: 10.3390/biology9100325
Tran HV, Nguyen QH. iAnt: combination of convolutional neural network and random forest models using PSSM and BERT features to identify antioxidant proteins. Curr Bioinform. 2022;17(2):184–95.
doi: 10.2174/1574893616666210820095144
Zhai Y, Zhang J, Zhang T, Gong Y, Zhang Z, Zhang D, Zhao Y. AOPM: application of antioxidant protein classification model in predicting the composition of antioxidant drugs. Front Pharmacol. 2022;12:818115.
pubmed: 35115948
pmcid: 8803896
doi: 10.3389/fphar.2021.818115
Meng C, Pei Y, Zou Q, Yuan L. DP-AOP: a novel SVM-based antioxidant proteins identifier. Int J Biol Macromol. 2023;247:125499.
pubmed: 37414318
doi: 10.1016/j.ijbiomac.2023.125499
Usman M, Khan S, Park S, Lee J-A. AoP-LSE: antioxidant proteins classification using deep latent space encoding of sequence features. Curr Issues Mol Biol. 2021;43(3):1489–501.
pubmed: 34698113
pmcid: 8928959
doi: 10.3390/cimb43030105
Ahmed S, Arif M, Kabir M, Khan K, Khan YD. PredAoDP: Accurate identification of antioxidant proteins by fusing different descriptors based on evolutionary information with support vector machine. Chemom Intell Lab Syst. 2022;228:104623.
doi: 10.1016/j.chemolab.2022.104623
Qin D, Jiao L, Wang R, Zhao Y, Hao Y, Liang G. Prediction of antioxidant peptides using a quantitative structure−activity relationship predictor (AnOxPP) based on bidirectional long short-term memory neural network and interpretable amino acid descriptors. Comput Biol Med. 2023;154:106591.
pubmed: 36701965
doi: 10.1016/j.compbiomed.2023.106591
Xi Q, Wang H, Yi L, Zhou J, Liang Y, Zhao X, Zuo Y. ANPrAod: identify antioxidant proteins by fusing amino acid clustering strategy and-peptide combination. Comput Math Methods Med. 2021;2021:1–10.
doi: 10.1155/2021/5518209
Olsen TH, Yesiltas B, Marin FI, Pertseva M, García-Moreno PJ, Gregersen S, Overgaard MT, Jacobsen C, Lund O, Hansen EB. AnOxPePred: using deep learning for the prediction of antioxidative properties of peptides. Sci Rep. 2020;10(1):21471.
pubmed: 33293615
pmcid: 7722737
doi: 10.1038/s41598-020-78319-w
Ahmad S, Charoenkwan P, Quinn JM, Moni MA, Hasan MM, Lio’ P, Shoombuatong W. SCORPION is a stacking-based ensemble learning framework for accurate prediction of phage virion proteins. Sci Rep. 2022;12(1):4106.
pubmed: 35260777
pmcid: 8904530
doi: 10.1038/s41598-022-08173-5
Chen Q, Wan Y, Lei Y, Zobel J, Verspoor K. Evaluation of CD-HIT for constructing non-redundant databases. In: 2016 IEEE international conference on bioinformatics and biomedicine (BIBM). IEEE; 2016. p. 703–6.
Ullah M, Akbar S, Raza A, Zou Q. DeepAVP-TPPred: identification of antiviral peptides using transformed image-based localized descriptors and binary tree growth algorithm. Bioinformatics. 2024;40:btae305.
pubmed: 38710482
pmcid: 11256913
doi: 10.1093/bioinformatics/btae305
Akbar S, Khan S, Ali F, Hayat M, Qasim M, Gul S. iHBP-DeepPSSM: identifying hormone binding proteins using PsePSSM based evolutionary features and deep learning approach. Chemom Intell Lab Syst. 2020;204:104103.
doi: 10.1016/j.chemolab.2020.104103
Yu B, Li S, Qiu W, Wang M, Du J, Zhang Y, Chen X. Prediction of subcellular location of apoptosis proteins by incorporating PsePSSM and DCCA coefficient based on LFDA dimensionality reduction. BMC Genom. 2018;19:1–17.
doi: 10.1186/s12864-018-4849-9
Nanni L, Brahnam S, Lumini A. Wavelet images and Chou’s pseudo amino acid composition for protein classification. Amino Acids. 2012;43:657–65.
pubmed: 21993538
doi: 10.1007/s00726-011-1114-9
Ahmad A, Akbar S, Tahir M, Hayat M, Ali F. iAFPs-EnC-GA: identifying antifungal peptides using sequential and evolutionary descriptors based multi-information fusion and ensemble learning approach. Chemom Intell Lab Syst. 2022;222:104516.
doi: 10.1016/j.chemolab.2022.104516
Lu W, Song Z, Ding Y, Wu H, Cao Y, Zhang Y, Li H. Use Chou’s 5-step rule to predict DNA-binding proteins with evolutionary information. BioMed Res Int. 2020;2020:1–9.
doi: 10.1155/2020/5204348
Zhang L, Zhao X, Kong L. Predict protein structural class for low-similarity sequences by evolutionary difference information into the general form of Chou׳ s pseudo amino acid composition. J Theor Biol. 2014;355:105–10.
pubmed: 24735902
doi: 10.1016/j.jtbi.2014.04.008
Sun D, Liu Z, Mao X, Yang Z, Ji C, Liu Y, Wang S. ANOX: a robust computational model for predicting the antioxidant proteins based on multiple features. Anal Biochem. 2021;631:114257.
pubmed: 34043981
doi: 10.1016/j.ab.2021.114257
Wang J, Yang B, Revote J, Leier A, Marquez-Lago TT, Webb G, Song J, Chou K-C, Lithgow T. POSSUM: a bioinformatics toolkit for generating numerical sequence feature descriptors based on PSSM profiles. Bioinformatics. 2017;33(17):2756–8.
pubmed: 28903538
doi: 10.1093/bioinformatics/btx302
Hayat M, Khan A. Mem-PHybrid: hybrid features-based prediction system for classifying membrane protein types. Anal Biochem. 2012;424(1):35–44.
pubmed: 22342883
doi: 10.1016/j.ab.2012.02.007
Hayat M, Khan A. WRF-TMH: predicting transmembrane helix by fusing composition index and physicochemical properties of amino acids. Amino Acids. 2013;44:1317–28.
pubmed: 23494269
doi: 10.1007/s00726-013-1466-4
Suvarna Vani K, Durga Bhavani S. SMOTE based protein fold prediction classification. In: Advances in computing and information technology: proceedings of the second international conference on advances in computing and information technology (ACITY) July 13–15, 2012, Chennai, India-Volume 2. Springer; 2013. p. 541–50.
Akbar S, Hayat M, Kabir M, Iqbal M. iAFP-gap-SMOTE: an efficient feature extraction scheme gapped dipeptide composition is coupled with an oversampling technique for identification of antifreeze proteins. Lett Org Chem. 2019;16(4):294–302.
doi: 10.2174/1570178615666180816101653
Hu J, He X, Yu D-J, Yang X-B, Yang J-Y, Shen H-B. A new supervised over-sampling algorithm with application to protein-nucleotide binding residue prediction. PLoS ONE. 2014;9(9):e107676.
pubmed: 25229688
pmcid: 4168127
doi: 10.1371/journal.pone.0107676
Elreedy D, Atiya AF. A comprehensive analysis of synthetic minority oversampling technique (SMOTE) for handling class imbalance. Inf Sci. 2019;505:32–64.
doi: 10.1016/j.ins.2019.07.070
Sun Y, Robinson M, Adams R, Te Boekhorst R, Rust AG, Davey N. Using sampling methods to improve binding site predictions. In: Proceedings of the 14th European symposium on artificial neural networks, ESANN 2006; 2006.
Chawla NV, Bowyer KW, Hall LO, Kegelmeyer WP. SMOTE: synthetic minority over-sampling technique. J Artif Intell Res. 2002;16:321–57.
doi: 10.1613/jair.953
Peng H, Long F, Ding C. Feature selection based on mutual information criteria of max-dependency, max-relevance, and min-redundancy. IEEE Trans Pattern Anal Mach Intell. 2005;27(8):1226–38.
pubmed: 16119262
doi: 10.1109/TPAMI.2005.159
Charoenkwan P, Chiangjong W, Nantasenamat C, Hasan MM, Manavalan B, Shoombuatong W. StackIL6: a stacking ensemble model for improving the prediction of IL-6 inducing peptides. Brief Bioinform. 2021;22(6):bbab172.
pubmed: 33963832
doi: 10.1093/bib/bbab172
Mishra A, Pokhrel P, Hoque MT. StackDPPred: a stacking based prediction of DNA-binding protein from sequence. Bioinformatics. 2019;35(3):433–41.
pubmed: 30032213
doi: 10.1093/bioinformatics/bty653
Basith S, Lee G, Manavalan B. STALLION: a stacking-based ensemble learning framework for prokaryotic lysine acetylation site prediction. Brief Bioinform. 2022;23(1):bbab376.
pubmed: 34532736
doi: 10.1093/bib/bbab376
Liang X, Li F, Chen J, Li J, Wu H, Li S, Song J, Liu Q. Large-scale comparative review and assessment of computational methods for anti-cancer peptide identification. Brief Bioinform. 2021;22(4):bbaa312.
pubmed: 33316035
doi: 10.1093/bib/bbaa312
Jiang M, Zhao B, Luo S, Wang Q, Chu Y, Chen T, Mao X, Liu Y, Wang Y, Jiang X. NeuroPpred-Fuse: an interpretable stacking model for prediction of neuropeptides by fusing sequence information and feature selection methods. Brief Bioinform. 2021;22(6):bbab310.
pubmed: 34396388
doi: 10.1093/bib/bbab310
Guo Y, Yan K, Lv H, Liu B. PreTP-EL: prediction of therapeutic peptides based on ensemble learning. Brief Bioinform. 2021;22(6):bbab358.
pubmed: 34459488
doi: 10.1093/bib/bbab358
Cao Z, Pan X, Yang Y, Huang Y, Shen H-B. The lncLocator: a subcellular localization predictor for long non-coding RNAs based on a stacked ensemble classifier. Bioinformatics. 2018;34(13):2185–94.
pubmed: 29462250
doi: 10.1093/bioinformatics/bty085
Zhang Q, Liu P, Wang X, Zhang Y, Han Y, Yu B. StackPDB: predicting DNA-binding proteins based on XGB-RFE feature optimization and stacked ensemble classifier. Appl Soft Comput. 2021;99:106921.
doi: 10.1016/j.asoc.2020.106921
Akbar S, Raza A, Zou Q. Deepstacked-AVPs: predicting antiviral peptides using tri-segment evolutionary profile and word embedding based multi-perspective features with deep stacking model. BMC Bioinform. 2024;25(1):102.
doi: 10.1186/s12859-024-05726-5
Akbar S, Ali H, Ahmad A, Sarker MR, Saeed A, Salwana E, Gul S, Khan A, Ali F. Prediction of amyloid proteins using embedded evolutionary & ensemble feature selection based descriptors with extreme gradient boosting model. IEEE Access; 2023.
Bukhari SNH, Webber J, Mehbodniya A. Decision tree based ensemble machine learning model for the prediction of Zika virus T-cell epitopes as potential vaccine candidates. Sci Rep. 2022;12(1):7810.
pubmed: 35552469
pmcid: 9096330
doi: 10.1038/s41598-022-11731-6
Akbar S, Rahman AU, Hayat M, Sohail M. cACP: classifying anticancer peptides using discriminative intelligent model via Chou’s 5-step rules and general pseudo components. Chemom Intell Lab Syst. 2020;196:103912.
doi: 10.1016/j.chemolab.2019.103912
Ao C, Zhou W, Gao L, Dong B, Yu L. Prediction of antioxidant proteins using hybrid feature representation method and random forest. Genomics. 2020;112(6):4666–74.
pubmed: 32818637
doi: 10.1016/j.ygeno.2020.08.016
Dwivedi AK. Performance evaluation of different machine learning techniques for prediction of heart disease. Neural Comput Appl. 2018;29:685–93.
doi: 10.1007/s00521-016-2604-1
Ali F, Akbar S, Ghulam A, Maher ZA, Unar A, Talpur DB. AFP-CMBPred: computational identification of antifreeze proteins by extending consensus sequences into multi-blocks evolutionary information. Comput Biol Med. 2021;139:105006.
pubmed: 34749096
doi: 10.1016/j.compbiomed.2021.105006
Akbar S, Zou Q, Raza A, Alarfaj FK. iAFPs-Mv-BiTCN: predicting antifungal peptides using self-attention transformer embedding and transform evolutionary based multi-view features with bidirectional temporal convolutional networks. Artif Intell Med. 2024;151:102860.
pubmed: 38552379
doi: 10.1016/j.artmed.2024.102860
Raza A, Uddin J, Almuhaimeed A, Akbar S, Zou Q, Ahmad A. AIPs-SnTCN: predicting anti-inflammatory peptides using fastText and transformer encoder-based hybrid word embedding with self-normalized temporal convolutional networks. J Chem Inf Model. 2023;63:6537–54.
pubmed: 37905969
doi: 10.1021/acs.jcim.3c01563
Raza A, Uddin J, Akbar S, Alarfaj FK, Zou Q, Ahmad A. Comprehensive analysis of computational methods for predicting anti-inflammatory peptides. Arch Comput Methods Eng. 2024. https://doi.org/10.1007/s11831-024-10078-7 .
doi: 10.1007/s11831-024-10078-7
Akbar S, Hayat M. iMethyl-STTNC: identification of N6-methyladenosine sites by extending the idea of SAAC into Chou’s PseAAC to formulate RNA sequences. J Theor Biol. 2018;455:205–11.
pubmed: 30031793
doi: 10.1016/j.jtbi.2018.07.018
Charoenkwan P, Ahmed S, Nantasenamat C, Quinn JM, Moni MA, Lio’ P, Shoombuatong W. AMYPred-FRL is a novel approach for accurate prediction of amyloid proteins by using feature representation learning. Sci Rep. 2022;12(1):7697.
pubmed: 35546347
pmcid: 9095707
doi: 10.1038/s41598-022-11897-z
Lundberg SM, Erion G, Chen H, DeGrave A, Prutkin JM, Nair B, Katz R, Himmelfarb J, Bansal N, Lee S-I. From local explanations to global understanding with explainable AI for trees. Nat Mach Intell. 2020;2(1):56–67.
pubmed: 32607472
pmcid: 7326367
doi: 10.1038/s42256-019-0138-9
Garreau D, Luxburg U. Explaining the explainer: a first theoretical analysis of LIME. In: International conference on artificial intelligence and statistics. PMLR; 2020. p. 1287–96.
Du Z, Ding X, Xu Y, Li Y. UniDL4BioPep: a universal deep learning architecture for binary classification in peptide bioactivity. Brief Bioinform. 2023;24(3):1–10.
doi: 10.1093/bib/bbad135