» Articles » PMID: 33671157

Combinatorial K-Means Clustering As a Machine Learning Tool Applied to Diabetes Mellitus Type 2

Overview
Publisher MDPI
Date 2021 Mar 6
PMID 33671157
Citations 10
Authors
Affiliations
Soon will be listed here.
Abstract

A new original procedure based on k-means clustering is designed to find the most appropriate clinical variables able to efficiently separate into groups similar patients diagnosed with diabetes mellitus type 2 (DMT2) and underlying diseases (arterial hypertonia (AH), ischemic heart disease (CHD), diabetic polyneuropathy (DPNP), and diabetic microangiopathy (DMA)). Clustering is a machine learning tool for discovering structures in datasets. Clustering has been proven to be efficient for pattern recognition based on clinical records. The considered combinatorial k-means procedure explores all possible k-means clustering with a determined number of descriptors and groups. The predetermined conditions for the partitioning were as follows: every single group of patients included patients with DMT2 and one of the underlying diseases; each subgroup formed in such a way was subject to partitioning into three patterns (good health status, medium health status, and degenerated health status); optimal descriptors for each disease and groups. The selection of the best clustering is obtained through the parameter called global variance, defined as the sum of all variance values of all clinical variables of all the clusters. The best clinical parameters are found by minimizing this global variance. This methodology has to identify a set of variables that are assumed to separate each underlying disease efficiently in three different subgroups of patients. The hierarchical clustering obtained for these four underlying diseases could be used to build groups of patients with correlated clinical data. The proposed methodology gives surmised results from complex data based on a relationship with the health status of the group and draws a picture of the prediction rate of the ongoing health status.

Citing Articles

Moving towards the use of artificial intelligence in pain management.

Antel R, Whitelaw S, Gore G, Ingelmo P Eur J Pain. 2024; 29(3):e4748.

PMID: 39523657 PMC: 11755729. DOI: 10.1002/ejp.4748.


The Use of Artificial Intelligence for Detecting and Predicting Atrial Arrhythmias Post Catheter Ablation.

Lallah P, Laite C, Bangash A, Chooah O, Jiang C Rev Cardiovasc Med. 2024; 24(8):215.

PMID: 39076714 PMC: 11266764. DOI: 10.31083/j.rcm2408215.


Unsupervised clustering analysis of comprehensive health status and its influencing factors on women of childbearing age: a cross-sectional study from a province in central China.

He L, Li S, Qin M, Yan Y, La Y, Cao X BMC Public Health. 2023; 23(1):2206.

PMID: 37946124 PMC: 10634171. DOI: 10.1186/s12889-023-17096-3.


A comprehensive review of machine learning algorithms and their application in geriatric medicine: present and future.

Woodman R, Mangoni A Aging Clin Exp Res. 2023; 35(11):2363-2397.

PMID: 37682491 PMC: 10627901. DOI: 10.1007/s40520-023-02552-2.


Identification of spinal tuberculosis subphenotypes using routine clinical data: a study based on unsupervised machine learning.

Yao Y, Wu S, Liu C, Zhou C, Zhu J, Chen T Ann Med. 2023; 55(2):2249004.

PMID: 37611242 PMC: 10448834. DOI: 10.1080/07853890.2023.2249004.


References
1.
Haq A, Li J, Khan J, Memon M, Nazir S, Ahmad S . Intelligent Machine Learning Approach for Effective Recognition of Diabetes in E-Healthcare Using Clinical Data. Sensors (Basel). 2020; 20(9). PMC: 7249007. DOI: 10.3390/s20092649. View

2.
Ryu K, Kang H, Lee S, Park H, You N, Kim J . Screening Model for Estimating Undiagnosed Diabetes among People with a Family History of Diabetes Mellitus: A KNHANES-Based Study. Int J Environ Res Public Health. 2020; 17(23). PMC: 7730533. DOI: 10.3390/ijerph17238903. View