» Articles » PMID: 15208204

Classification of Gene Microarrays by Penalized Logistic Regression

Overview
Journal Biostatistics
Specialty Public Health
Date 2004 Jun 23
PMID 15208204
Citations 72
Authors
Affiliations
Soon will be listed here.
Abstract

Classification of patient samples is an important aspect of cancer diagnosis and treatment. The support vector machine (SVM) has been successfully applied to microarray cancer diagnosis problems. However, one weakness of the SVM is that given a tumor sample, it only predicts a cancer class label but does not provide any estimate of the underlying probability. We propose penalized logistic regression (PLR) as an alternative to the SVM for the microarray cancer diagnosis problem. We show that when using the same set of genes, PLR and the SVM perform similarly in cancer classification, but PLR has the advantage of additionally providing an estimate of the underlying probability. Often a primary goal in microarray cancer diagnosis is to identify the genes responsible for the classification, rather than class prediction. We consider two gene selection methods in this paper, univariate ranking (UR) and recursive feature elimination (RFE). Empirical results indicate that PLR combined with RFE tends to select fewer genes than other methods and also performs well in both cross-validation and test samples. A fast algorithm for solving PLR is also described.

Citing Articles

Ensemble Classification Model With CFS-IGWO-Based Feature Selection for Cancer Detection Using Microarray Data.

Panda P, Bisoy S, Kautish S, Ahmad R, Irshad A, Sarwar N Int J Telemed Appl. 2024; 2024:4105224.

PMID: 39449963 PMC: 11502127. DOI: 10.1155/2024/4105224.


Low-rank regression models for multiple binary responses and their applications to cancer cell-line encyclopedia data.

Park S, Lee E, Zhao H J Am Stat Assoc. 2024; 119(545):202-216.

PMID: 38481466 PMC: 10928550. DOI: 10.1080/01621459.2022.2105704.


Machine learning algorithms for identifying predictive variables of mortality risk following dementia diagnosis: a longitudinal cohort study.

Mostafaei S, Hoang M, Jurado P, Xu H, Zacarias-Pons L, Eriksdotter M Sci Rep. 2023; 13(1):9480.

PMID: 37301891 PMC: 10257644. DOI: 10.1038/s41598-023-36362-3.


Identification of miRNA biomarkers for breast cancer by combining ensemble regularized multinomial logistic regression and Cox regression.

Li J, Zhang H, Gao F BMC Bioinformatics. 2022; 23(1):434.

PMID: 36258162 PMC: 9580207. DOI: 10.1186/s12859-022-04982-7.


Classification of COVID19 Patients Using Robust Logistic Regression.

Ghosh A, Jaenada M, Pardo L J Stat Theory Pract. 2022; 16(4):67.

PMID: 36164412 PMC: 9491676. DOI: 10.1007/s42519-022-00295-3.