Score: 0

Scalable Medication Extraction and Discontinuation Identification from Electronic Health Records Using Large Language Models

Published: June 10, 2025 | arXiv ID: 2506.11137v1

By: Chong Shao , Douglas Snyder , Chiran Li and more

Potential Business Impact:

Helps doctors find forgotten medicines in patient notes.

Business Areas:

Natural Language Processing Artificial Intelligence, Data and Analytics, Software

Identifying medication discontinuations in electronic health records (EHRs) is vital for patient safety but is often hindered by information being buried in unstructured notes. This study aims to evaluate the capabilities of advanced open-sourced and proprietary large language models (LLMs) in extracting medications and classifying their medication status from EHR notes, focusing on their scalability on medication information extraction without human annotation. We collected three EHR datasets from diverse sources to build the evaluation benchmark. We evaluated 12 advanced LLMs and explored multiple LLM prompting strategies. Performance on medication extraction, medication status classification, and their joint task (extraction then classification) was systematically compared across all experiments. We found that LLMs showed promising performance on the medication extraction and discontinuation classification from EHR notes. GPT-4o consistently achieved the highest average F1 scores in all tasks under zero-shot setting - 94.0% for medication extraction, 78.1% for discontinuation classification, and 72.7% for the joint task. Open-sourced models followed closely, Llama-3.1-70B-Instruct achieved the highest performance in medication status classification on the MIV-Med dataset (68.7%) and in the joint task on both the Re-CASI (76.2%) and MIV-Med (60.2%) datasets. Medical-specific LLMs demonstrated lower performance compared to advanced general-domain LLMs. Few-shot learning generally improved performance, while CoT reasoning showed inconsistent gains. LLMs demonstrate strong potential for medication extraction and discontinuation identification on EHR notes, with open-sourced models offering scalable alternatives to proprietary systems and few-shot can further improve LLMs' capability.

Integrating Large Language Models with Human Expertise for Disease Detection in Electronic Health Records

Computation and Language

Helps doctors find patient sicknesses faster.

31 Mar 2025 0

89%

Large Language Models for Drug Overdose Prediction from Longitudinal Medical Records

Artificial Intelligence

Helps doctors predict overdose risk from patient records.

16 Apr 2025 0

89%

Benchmarking Open-Source Large Language Models on Healthcare Text Classification Tasks

Computation and Language

Helps computers find health info from text.

19 Mar 2025 1

View PDF Login to Bookmark

Country of Origin

🇺🇸 United States

Page Count

35 pages

Scalable Medication Extraction and Discontinuation Identification from Electronic Health Records Using Large Language Models

Helps doctors find forgotten medicines in patient notes.

Technical Abstract

Integrating Large Language Models with Human Expertise for Disease Detection in Electronic Health Records

Large Language Models for Drug Overdose Prediction from Longitudinal Medical Records

Benchmarking Open-Source Large Language Models on Healthcare Text Classification Tasks