» Articles » PMID: 30415424

Detecting Significant Genotype-phenotype Association Rules in Bipolar Disorder: Market Research Meets Complex Genetics

Abstract

Background: Disentangling the etiology of common, complex diseases is a major challenge in genetic research. For bipolar disorder (BD), several genome-wide association studies (GWAS) have been performed. Similar to other complex disorders, major breakthroughs in explaining the high heritability of BD through GWAS have remained elusive. To overcome this dilemma, genetic research into BD, has embraced a variety of strategies such as the formation of large consortia to increase sample size and sequencing approaches. Here we advocate a complementary approach making use of already existing GWAS data: a novel data mining procedure to identify yet undetected genotype-phenotype relationships. We adapted association rule mining, a data mining technique traditionally used in retail market research, to identify frequent and characteristic genotype patterns showing strong associations to phenotype clusters. We applied this strategy to three independent GWAS datasets from 2835 phenotypically characterized patients with BD. In a discovery step, 20,882 candidate association rules were extracted.

Results: Two of these rules-one associated with eating disorder and the other with anxiety-remained significant in an independent dataset after robust correction for multiple testing. Both showed considerable effect sizes (odds ratio ~ 3.4 and 3.0, respectively) and support previously reported molecular biological findings.

Conclusion: Our approach detected novel specific genotype-phenotype relationships in BD that were missed by standard analyses like GWAS. While we developed and applied our method within the context of BD gene discovery, it may facilitate identifying highly specific genotype-phenotype relationships in subsets of genome-wide data sets of other complex phenotype with similar epidemiological properties and challenges to gene discovery efforts.

Citing Articles

Dopaminergic Epistases in Schizophrenia.

Bosun A, Albu-Kalinovic R, Neda-Stepan O, Bosun I, Farcas S, Enatescu V Brain Sci. 2024; 14(11).

PMID: 39595853 PMC: 11592377. DOI: 10.3390/brainsci14111089.


First-in-human, double-blind, randomized phase 1b study of peptide immunotherapy IMCY-0098 in new-onset type 1 diabetes: an exploratory analysis of immune biomarkers.

van Rampelbergh J, Achenbach P, Leslie R, Kindermans M, Parmentier F, Carlier V BMC Med. 2024; 22(1):259.

PMID: 38902652 PMC: 11191262. DOI: 10.1186/s12916-024-03476-y.


The diabetes mellitus multimorbidity network in hospitalized patients over 50 years of age in China: data mining of medical records.

Chen C, Zheng X, Liao S, Chen S, Liang M, Tang K BMC Public Health. 2024; 24(1):1433.

PMID: 38811975 PMC: 11134652. DOI: 10.1186/s12889-024-18887-y.


Prevalence of eating disorders in patients with bipolar disorder: a scoping review of the literature.

Yakovleva Y, Kasyanov E, Mazo G Consort Psychiatr. 2024; 4(2):91-106.

PMID: 38250644 PMC: 10795952. DOI: 10.17816/CP6338.


Characteristics of insulin resistance in Korean adults from the perspective of circadian and metabolic sensing genes.

Park M, Lee S, Baek Y, Lee J, Park S, Cho J Genes Genomics. 2023; 45(12):1475-1487.

PMID: 37768516 PMC: 10682234. DOI: 10.1007/s13258-023-01443-0.


References
1.
McMahon F, Akula N, Schulze T, Muglia P, Tozzi F, Detera-Wadleigh S . Meta-analysis of genome-wide association data identifies a risk locus for major mood disorders on 3p21.1. Nat Genet. 2010; 42(2):128-31. PMC: 2854040. DOI: 10.1038/ng.523. View

2.
Benjamini Y . Simultaneous and selective inference: Current successes and future challenges. Biom J. 2010; 52(6):708-21. DOI: 10.1002/bimj.200900299. View

3.
Biel M, Seeliger M, Pfeifer A, Kohler K, Gerstner A, Ludwig A . Selective loss of cone function in mice lacking the cyclic nucleotide-gated channel CNG3. Proc Natl Acad Sci U S A. 1999; 96(13):7553-7. PMC: 22124. DOI: 10.1073/pnas.96.13.7553. View

4.
Lee K, Woon P, Teo Y, Sim K . Genome wide association studies (GWAS) and copy number variation (CNV) studies of the major psychoses: what have we learnt?. Neurosci Biobehav Rev. 2011; 36(1):556-71. DOI: 10.1016/j.neubiorev.2011.09.001. View

5.
. Large-scale genome-wide association analysis of bipolar disorder identifies a new susceptibility locus near ODZ4. Nat Genet. 2011; 43(10):977-83. PMC: 3637176. DOI: 10.1038/ng.943. View