CARE-SD: Classifier-based Analysis for Recognizing Provider Stigmatizing and Doubt Marker Labels in Electronic Health Records: Model Development and Validation

Overview

Journal J Am Med Inform Assoc

Publisher Oxford University Press

Specialty Medical Informatics

Date 2024 Dec 26

PMID 39724920

Authors

Andrew Walker

Annie Thorne

Sudeshna Das

Jennifer Love

Hannah L F Cooper

Melvin Livingston 3rd

Abeed Sarker

Affiliations

Soon will be listed here.

Abstract

Objective: To detect and classify features of stigmatizing and biased language in intensive care electronic health records (EHRs) using natural language processing techniques.

Materials And Methods: We first created a lexicon and regular expression lists from literature-driven stem words for linguistic features of stigmatizing patient labels, doubt markers, and scare quotes within EHRs. The lexicon was further extended using Word2Vec and GPT 3.5, and refined through human evaluation. These lexicons were used to search for matches across 18 million sentences from the de-identified Medical Information Mart for Intensive Care-III (MIMIC-III) dataset. For each linguistic bias feature, 1000 sentence matches were sampled, labeled by expert clinical and public health annotators, and used to supervised learning classifiers.

Results: Lexicon development from expanded literature stem-word lists resulted in a doubt marker lexicon containing 58 expressions, and a stigmatizing labels lexicon containing 127 expressions. Classifiers for doubt markers and stigmatizing labels had the highest performance, with macro F1-scores of 0.84 and 0.79, positive-label recall and precision values ranging from 0.71 to 0.86, and accuracies aligning closely with human annotator agreement (0.87).

Discussion: This study demonstrated the feasibility of supervised classifiers in automatically identifying stigmatizing labels and doubt markers in medical text and identified trends in stigmatizing language use in an EHR setting. Additional labeled data may help improve lower scare quote model performance.

Conclusions: Classifiers developed in this study showed high model performance and can be applied to identify patterns and target interventions to reduce stigmatizing labels and doubt markers in healthcare systems.

References

Nembrini S, Konig I, Wright M . The revival of the Gini importance?. Bioinformatics. 2018; 34(21):3711-3718. PMC: 6198850. DOI: 10.1093/bioinformatics/bty373. View

Chen L, Crum R, Martins S, Kaufmann C, Strain E, Mojtabai R . Service use and barriers to mental health care among adults with major depression and comorbid substance dependence. Psychiatr Serv. 2013; 64(9):863-70. PMC: 4049190. DOI: 10.1176/appi.ps.201200289. View

FitzGerald C, Hurst S . Implicit bias in healthcare professionals: a systematic review. BMC Med Ethics. 2017; 18(1):19. PMC: 5333436. DOI: 10.1186/s12910-017-0179-8. View

Hatzenbuehler M, Phelan J, Link B . Stigma as a fundamental cause of population health inequalities. Am J Public Health. 2013; 103(5):813-21. PMC: 3682466. DOI: 10.2105/AJPH.2012.301069. View

Beach M, Saha S . Quoting Patients in Clinical Notes: First, Do No Harm. Ann Intern Med. 2021; 174(10):1454-1455. DOI: 10.7326/M21-2449. View

Himmelstein G, Bates D, Zhou L . Examination of Stigmatizing Language in the Electronic Health Record. JAMA Netw Open. 2022; 5(1):e2144967. PMC: 8796019. DOI: 10.1001/jamanetworkopen.2021.44967. View

Zestcott C, Spece L, McDermott D, Stone J . Health Care Providers' Negative Implicit Attitudes and Stereotypes of American Indians. J Racial Ethn Health Disparities. 2020; 8(1):230-236. DOI: 10.1007/s40615-020-00776-w. View

Park J, Saha S, Chee B, Taylor J, Beach M . Physician Use of Stigmatizing Language in Patient Medical Records. JAMA Netw Open. 2021; 4(7):e2117052. PMC: 8281008. DOI: 10.1001/jamanetworkopen.2021.17052. View

Barcelona V, Scharp D, Moen H, Davoudi A, Idnay B, Cato K . Using Natural Language Processing to Identify Stigmatizing Language in Labor and Birth Clinical Notes. Matern Child Health J. 2023; 28(3):578-586. DOI: 10.1007/s10995-023-03857-4. View

10.

Maina I, Belton T, Ginzberg S, Singh A, Johnson T . A decade of studying implicit racial/ethnic bias in healthcare providers using the implicit association test. Soc Sci Med. 2017; 199:219-229. DOI: 10.1016/j.socscimed.2017.05.009. View

11.

DesRoches C . Healthcare in the new age of transparency. Semin Dial. 2020; 33(6):533-538. DOI: 10.1111/sdi.12934. View

12.

Johnson A, Pollard T, Shen L, Lehman L, Feng M, Ghassemi M . MIMIC-III, a freely accessible critical care database. Sci Data. 2016; 3:160035. PMC: 4878278. DOI: 10.1038/sdata.2016.35. View

13.

Sun M, Oliwa T, Peek M, Tung E . Negative Patient Descriptors: Documenting Racial Bias In The Electronic Health Record. Health Aff (Millwood). 2022; 41(2):203-211. PMC: 8973827. DOI: 10.1377/hlthaff.2021.01423. View

14.

Ross L, Vigod S, Wishart J, Waese M, Spence J, Oliver J . Barriers and facilitators to primary care for people with mental health and/or substance use issues: a qualitative study. BMC Fam Pract. 2015; 16:135. PMC: 4604001. DOI: 10.1186/s12875-015-0353-3. View

15.

Rodriguez J, Clark C, Bates D . Digital Health Equity as a Necessity in the 21st Century Cures Act Era. JAMA. 2020; 323(23):2381-2382. DOI: 10.1001/jama.2020.7858. View

16.

Zhang Y, Chen Q, Yang Z, Lin H, Lu Z . BioWordVec, improving biomedical word embeddings with subword information and MeSH. Sci Data. 2019; 6(1):52. PMC: 6510737. DOI: 10.1038/s41597-019-0055-0. View

17.

Bremer W, Plaisance K, Walker D, Bonn M, Love J, Perrone J . Barriers to opioid use disorder treatment: A comparison of self-reported information from social media with barriers found in literature. Front Public Health. 2023; 11:1141093. PMC: 10158842. DOI: 10.3389/fpubh.2023.1141093. View

18.

Beach M, Saha S, Park J, Taylor J, Drew P, Plank E . Testimonial Injustice: Linguistic Bias in the Medical Records of Black Patients and Women. J Gen Intern Med. 2021; 36(6):1708-1714. PMC: 8175470. DOI: 10.1007/s11606-021-06682-z. View

19.

Link B, Phelan J . Stigma and its public health implications. Lancet. 2006; 367(9509):528-9. DOI: 10.1016/S0140-6736(06)68184-1. View

20.

Goddu A, OConor K, Lanzkron S, Saheed M, Saha S, Peek M . Do Words Matter? Stigmatizing Language and the Transmission of Bias in the Medical Record. J Gen Intern Med. 2018; 33(5):685-691. PMC: 5910343. DOI: 10.1007/s11606-017-4289-2. View