» Articles » PMID: 29297322

CNN-BLPred: a Convolutional Neural Network Based Predictor for β-Lactamases (BL) and Their Classes

Overview
Publisher Biomed Central
Specialty Biology
Date 2018 Jan 4
PMID 29297322
Citations 13
Authors
Affiliations
Soon will be listed here.
Abstract

Background: The β-Lactamase (BL) enzyme family is an important class of enzymes that plays a key role in bacterial resistance to antibiotics. As the newly identified number of BL enzymes is increasing daily, it is imperative to develop a computational tool to classify the newly identified BL enzymes into one of its classes. There are two types of classification of BL enzymes: Molecular Classification and Functional Classification. Existing computational methods only address Molecular Classification and the performance of these existing methods is unsatisfactory.

Results: We addressed the unsatisfactory performance of the existing methods by implementing a Deep Learning approach called Convolutional Neural Network (CNN). We developed CNN-BLPred, an approach for the classification of BL proteins. The CNN-BLPred uses Gradient Boosted Feature Selection (GBFS) in order to select the ideal feature set for each BL classification. Based on the rigorous benchmarking of CCN-BLPred using both leave-one-out cross-validation and independent test sets, CCN-BLPred performed better than the other existing algorithms. Compared with other architectures of CNN, Recurrent Neural Network, and Random Forest, the simple CNN architecture with only one convolutional layer performs the best. After feature extraction, we were able to remove ~95% of the 10,912 features using Gradient Boosted Trees. During 10-fold cross validation, we increased the accuracy of the classic BL predictions by 7%. We also increased the accuracy of Class A, Class B, Class C, and Class D performance by an average of 25.64%. The independent test results followed a similar trend.

Conclusions: We implemented a deep learning algorithm known as Convolutional Neural Network (CNN) to develop a classifier for BL classification. Combined with feature selection on an exhaustive feature set and using balancing method such as Random Oversampling (ROS), Random Undersampling (RUS) and Synthetic Minority Oversampling Technique (SMOTE), CNN-BLPred performs significantly better than existing algorithms for BL classification.

Citing Articles

An endoscopic ultrasound-based interpretable deep learning model and nomogram for distinguishing pancreatic neuroendocrine tumors from pancreatic cancer.

Yi N, Mo S, Zhang Y, Jiang Q, Wang Y, Huang C Sci Rep. 2025; 15(1):3383.

PMID: 39870667 PMC: 11772604. DOI: 10.1038/s41598-024-84749-7.


Predictors of the rate of cognitive decline in older adults using machine learning.

Ahmadzadeh M, Cosco T, Best J, Christie G, DiPaola S PLoS One. 2023; 18(3):e0280029.

PMID: 36867596 PMC: 9983884. DOI: 10.1371/journal.pone.0280029.


Factors related to steroid treatment responsiveness in thyroid eye disease patients and application of SHAP for feature analysis with XGBoost.

Park J, Kim J, Ryu D, Choi H Front Endocrinol (Lausanne). 2023; 14:1079628.

PMID: 36817584 PMC: 9928572. DOI: 10.3389/fendo.2023.1079628.


β-LacFamPred: An online tool for prediction and classification of β-lactamase class, subclass, and family.

Pandey D, Singhal N, Kumar M Front Microbiol. 2023; 13:1039687.

PMID: 36713195 PMC: 9878453. DOI: 10.3389/fmicb.2022.1039687.


Identification of the ubiquitin-proteasome pathway domain by hyperparameter optimization based on a 2D convolutional neural network.

Sikander R, Arif M, Ghulam A, Worachartcheewan A, A Thafar M, Habib S Front Genet. 2022; 13:851688.

PMID: 35937990 PMC: 9355632. DOI: 10.3389/fgene.2022.851688.


References
1.
Li W, Godzik A . Cd-hit: a fast program for clustering and comparing large sets of protein or nucleotide sequences. Bioinformatics. 2006; 22(13):1658-9. DOI: 10.1093/bioinformatics/btl158. View

2.
Lu P, Hsieh Y, Lin J, Huang J, Yang T, Lin L . Characterisation of fosfomycin resistance mechanisms and molecular epidemiology in extended-spectrum β-lactamase-producing Klebsiella pneumoniae isolates. Int J Antimicrob Agents. 2016; 48(5):564-568. DOI: 10.1016/j.ijantimicag.2016.08.013. View

3.
Ismail H, Newman R, Kc D . RF-Hydroxysite: a random forest based predictor for hydroxylation sites. Mol Biosyst. 2016; 12(8):2427-35. PMC: 4955772. DOI: 10.1039/c6mb00179c. View

4.
LeCun Y, Bengio Y, Hinton G . Deep learning. Nature. 2015; 521(7553):436-44. DOI: 10.1038/nature14539. View

5.
Thai Q, Pleiss J . SHV Lactamase Engineering Database: a reconciliation tool for SHV β-lactamases in public databases. BMC Genomics. 2010; 11:563. PMC: 3091712. DOI: 10.1186/1471-2164-11-563. View