» Articles » PMID: 39387466

Clustering-based Risk Stratification of Prediabetes Populations: Insights from the Taiwan and UK Biobanks

Overview
Specialty Endocrinology
Date 2024 Oct 10
PMID 39387466
Authors
Affiliations
Soon will be listed here.
Abstract

Aims/introduction: This study aimed to identify low- and high-risk diabetes groups within prediabetes populations using data from the Taiwan Biobank (TWB) and UK Biobank (UKB) through a clustering-based Unsupervised Learning (UL) approach, to inform targeted type 2 diabetes (T2D) interventions.

Materials And Methods: Data from TWB and UKB, comprising clinical and genetic information, were analyzed. Prediabetes was defined by glucose thresholds, and incident T2D was identified through follow-up data. K-means clustering was performed on prediabetes participants using significant features determined through logistic regression and LASSO. Cluster stability was assessed using mean Jaccard similarity, silhouette score, and the elbow method.

Results: We identified two stable clusters representing high- and low-risk diabetes groups in both biobanks. The high-risk clusters showed higher diabetes incidence, with 15.7% in TWB and 13.0% in UKB, compared to 7.3% and 9.1% in the low-risk clusters, respectively. Notably, males were predominant in the high-risk groups, constituting 76.6% in TWB and 52.7% in UKB. In TWB, the high-risk group also exhibited significantly higher BMI, fasting glucose, and triglycerides, while UKB showed marginal significance in BMI and other metabolic indicators. Current smoking was significantly associated with increased diabetes risk in the TWB high-risk group (P < 0.001). Kaplan-Meier curves indicated significant differences in diabetes complication incidences between clusters.

Conclusions: UL effectively identified risk-specific groups within prediabetes populations, with high-risk groups strongly associated male gender, higher BMI, smoking, and metabolic markers. Tailored preventive strategies, particularly for young males in Taiwan, are crucial to reducing T2D risk.

Citing Articles

Clinical symptoms and functional impairment in attention deficit hyperactivity disorder (ADHD) co-morbid tic disorder (TD) patients: a cluster-based investigation.

Jiang Z, Xu H, Zhang A, Yu L, Wang X, Zhang W BMC Psychiatry. 2025; 25(1):100.

PMID: 39905386 PMC: 11796001. DOI: 10.1186/s12888-025-06558-0.

References
1.
Chung R, Chuang S, Chen Y, Li G, Hsieh C, Chiou H . Prevalence and predictive modeling of undiagnosed diabetes and impaired fasting glucose in Taiwan: a Taiwan Biobank study. BMJ Open Diabetes Res Care. 2023; 11(3). PMC: 10277095. DOI: 10.1136/bmjdrc-2023-003423. View

2.
Barbu E, Popescu M, Popescu A, Balanescu S . Phenotyping the Prediabetic Population-A Closer Look at Intermediate Glucose Status and Cardiovascular Disease. Int J Mol Sci. 2021; 22(13). PMC: 8268766. DOI: 10.3390/ijms22136864. View

3.
Lin C, Lo F, Huang Y, Sun J, Chen S, Kuo C . Evaluation of Disease Complications Among Adults With Type 1 Diabetes and a Family History of Type 2 Diabetes in Taiwan. JAMA Netw Open. 2021; 4(12):e2138775. PMC: 8672229. DOI: 10.1001/jamanetworkopen.2021.38775. View

4.
Liu C, Chang C, Chen I, Lin F, Tzou S, Hsieh C . Machine Learning Prediction of Prediabetes in a Young Male Chinese Cohort with 5.8-Year Follow-Up. Diagnostics (Basel). 2024; 14(10). PMC: 11119884. DOI: 10.3390/diagnostics14100979. View

5.
Southern D, Norris C, Quan H, Shrive F, Galbraith P, Humphries K . An administrative data merging solution for dealing with missing data in a clinical registry: adaptation from ICD-9 to ICD-10. BMC Med Res Methodol. 2008; 8:1. PMC: 2244639. DOI: 10.1186/1471-2288-8-1. View