Extracting Patient History from Clinical Text: A Comparative Study of Clinical Large Language Models

๐Ÿ“… 2025-03-30
๐Ÿ“ˆ Citations: 0
โœจ Influential: 0
๐Ÿ“„ PDF
๐Ÿค– AI Summary
This study addresses the automated extraction of Medical History Entities (MHEs)โ€”including Chief Complaint (CC), History of Present Illness (HPI), and Past/Family/Social History (PFSH)โ€”from clinical narratives to improve the conversion of unstructured electronic health records (EHRs) into standardized formats. Methodologically, it conducts the first systematic evaluation of seven clinical large language models (cLLMs), including fine-tuned GatorTron/GatorTronS and zero-shot GPT-4o, using a fine-grained manually annotated MTSamples dataset; error analysis and ablation studies examine impacts of text segmentation, entity length, and other linguistic features. A novel contribution is the integration of foundational medical entities (BMEs) as auxiliary signals. Results show that fine-tuned cLLMs reduce MHE extraction latency by over 20%; GatorTron variants achieve the highest performance; BME augmentation improves F1 scores for certain MHE types by up to 5.3%; and explicitly structured, title-annotated paragraphs significantly enhance extraction accuracy.

Technology Category

Natural Language Processing: Information ExtractionMachine Learning: Large Multimodal Models (LMMs)Data Mining & Knowledge Management: Conversational Systems for Recommendation & Retrieval

Application Category

Search and Retrieval-Augmented AI: Large language models for searchWeb Mining and Content Analysis: Large pretrained models with web dataSemantics and Knowledge: Methods to enhance, augment, integrate or synergize semantic models such as knowledge graphs and LLMs
๐Ÿ“ Abstract
Extracting medical history entities (MHEs) related to a patient's chief complaint (CC), history of present illness (HPI), and past, family, and social history (PFSH) helps structure free-text clinical notes into standardized EHRs, streamlining downstream tasks like continuity of care, medical coding, and quality metrics. Fine-tuned clinical large language models (cLLMs) can assist in this process while ensuring the protection of sensitive data via on-premises deployment. This study evaluates the performance of cLLMs in recognizing CC/HPI/PFSH-related MHEs and examines how note characteristics impact model accuracy. We annotated 1,449 MHEs across 61 outpatient-related clinical notes from the MTSamples repository. To recognize these entities, we fine-tuned seven state-of-the-art cLLMs. Additionally, we assessed the models' performance when enhanced by integrating, problems, tests, treatments, and other basic medical entities (BMEs). We compared the performance of these models against GPT-4o in a zero-shot setting. To further understand the textual characteristics affecting model accuracy, we conducted an error analysis focused on note length, entity length, and segmentation. The cLLMs showed potential in reducing the time required for extracting MHEs by over 20%. However, detecting many types of MHEs remained challenging due to their polysemous nature and the frequent involvement of non-medical vocabulary. Fine-tuned GatorTron and GatorTronS, two of the most extensively trained cLLMs, demonstrated the highest performance. Integrating pre-identified BME information improved model performance for certain entities. Regarding the impact of textual characteristics on model performance, we found that longer entities were harder to identify, note length did not correlate with a higher error rate, and well-organized segments with headings are beneficial for the extraction.
Problem

Research questions and friction points this paper is trying to address.

Evaluating clinical LLMs for extracting medical history entities from notes
Assessing impact of note characteristics on entity recognition accuracy
Comparing fine-tuned cLLMs against GPT-4o in zero-shot settings
Innovation

Methods, ideas, or system contributions that make the work stand out.

Fine-tuned clinical LLMs for medical history extraction
On-premises deployment ensures sensitive data protection
Integration of basic medical entities boosts performance
๐Ÿ”Ž Similar Papers
No similar papers found.
๐Ÿ’ผ Related Jobs
No related jobs found.
H
Hieu Nghiem
Center for Health Systems Innovation, Oklahoma State University, Stillwater, OK, 74078, USA
T
Tuan-Dung Le
Department of Machine Learning, Moffitt Cancer Center and Research Institute, Tampa, FL, 33612, USA
S
Suhao Chen
Department of Industrial Engineering, South Dakota School of Mines and Technology, Rapid City, SD, 57701, USA
T
Thanh Thieu
Department of Machine Learning, Moffitt Cancer Center and Research Institute, Tampa, FL, 33612, USA; Department of Oncological Sciences, University of South Florida Morsani College of Medicine, Tampa, FL, 33612, USA
A
Andrew Gin
Center for Health Systems Innovation, Oklahoma State University, Stillwater, OK, 74078, USA
E
Ellie Phuong Nguyen
South University School of Pharmacy, Savannah, GA, 31406, USA
Dursun Delen
Dursun Delen
Regents Professor, Spears and Patterson Chairs, Spears School of Business, Oklahoma State University
Health AnalyticsDecision Support SystemsHealthcare AnalyticsBusiness IntelligenceBusiness Analytics
J
Johnson Thomas
Department of Computer Science, Oklahoma State University, Stillwater, OK, 74078, USA
J
Jivan Lamichhane
Department of Medicine, The State University of New York Upstate Medical University, Syracuse, NY, 13210, USA
Z
Zhuqi Miao
School of Business, The State University of New York at New Paltz, New Paltz, NY, 12561, USA