» Articles » PMID: 19646551

Rule-based Information Extraction from Patients' Clinical Data

Overview
Journal J Biomed Inform
Publisher Elsevier
Date 2009 Aug 4
PMID 19646551
Citations 33
Authors
Affiliations
Soon will be listed here.
Abstract

The paper describes a rule-based information extraction (IE) system developed for Polish medical texts. We present two applications designed to select data from medical documentation in Polish: mammography reports and hospital records of diabetic patients. First, we have designed a special ontology that subsequently had its concepts translated into two separate models, represented as typed feature structure (TFS) hierarchies, complying with the format required by the IE platform we adopted. Then, we used dedicated IE grammars to process documents and fill in templates provided by the models. In particular, in the grammars, we addressed such linguistic issues as: ambiguous keywords, negation, coordination or anaphoric expressions. Resolving some of these problems has been deferred to a post-processing phase where the extracted information is further grouped and structured into more complex templates. To this end, we defined special heuristic algorithms on the basis of sample data. The evaluation of the implemented procedures shows their usability for clinical data extraction tasks. For most of the evaluated templates, precision and recall well above 80% were obtained.

Citing Articles

Attention-based interactive multi-level feature fusion for named entity recognition.

Xu Y, Chen Y Sci Rep. 2025; 15(1):3069.

PMID: 39856193 PMC: 11760954. DOI: 10.1038/s41598-025-86718-0.


Automated Transformation of Unstructured Cardiovascular Diagnostic Reports into Structured Datasets Using Sequentially Deployed Large Language Models.

Vasisht Shankar S, Dhingra L, Aminorroaya A, Adejumo P, Nadkarni G, Xu H medRxiv. 2024; .

PMID: 39417094 PMC: 11482995. DOI: 10.1101/2024.10.08.24315035.


Comparison of Diagnosis Codes to Clinical Notes in Classifying Patients with Diabetic Retinopathy.

Yonamine S, Ma C, Alabi R, Kaidonis G, Chan L, Borkar D Ophthalmol Sci. 2024; 4(6):100564.

PMID: 39253554 PMC: 11382306. DOI: 10.1016/j.xops.2024.100564.


Use of Generative AI to Identify Helmet Status Among Patients With Micromobility-Related Injuries From Unstructured Clinical Notes.

Burford K, Itzkowitz N, Ortega A, Teitler J, Rundle A JAMA Netw Open. 2024; 7(8):e2425981.

PMID: 39136946 PMC: 11322845. DOI: 10.1001/jamanetworkopen.2024.25981.


An Empirical Evaluation of Prompting Strategies for Large Language Models in Zero-Shot Clinical Natural Language Processing: Algorithm Development and Validation Study.

Sivarajkumar S, Kelley M, Samolyk-Mazzanti A, Visweswaran S, Wang Y JMIR Med Inform. 2024; 12:e55318.

PMID: 38587879 PMC: 11036183. DOI: 10.2196/55318.