PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
October 23, 2025JAMA Network Open29 citationsOpen Access

Large Language Model Assistant for Emergency Department Discharge Documentation

View Full Paper
JSJi Woo SongJPJunseong ParkJKJi Hoon Kim

Key Points

  • Emergency department discharge notes saw improved conciseness with LLM assistance, achieving higher effectiveness.
  • Mean documentation time per note dropped from 69.5 seconds for manual notes to 32.0 seconds with LLM support.
  • Comparative effectiveness study utilized patient records from a 2400-bed tertiary care hospital for validation.
  • Findings suggest significant workflow efficiency through LLM integration in emergency documentation processes.

Abstract

Importance Emergency department (ED) discharge documentation is time-consuming and often incomplete. Objective To develop a large language model (LLM) assistant that generates ED discharge notes and to evaluate its effectiveness on documentation quality and workflow efficiency. Design, Setting, and Participants This comparative effectiveness study, which was conducted at a 2400-bed tertiary care hospital in South Korea, consisted of 2 primary phases: a development phase and sequential validation of the LLM assistant. In the randomized sequential prospective validation, 6 emergency physicians first wrote discharge notes manually (session 1), then edited LLM-generated drafts after a 1-hour washout period (session 2). Three independent physicians evaluated 300 note sets (each containing a manual note, an LLM draft, and an LLM-assisted note). For model development and validation, patient records from ED visits between September 1, 2022, and August 31, 2023, were used. The inclusion criteria encompassed adult patients (aged ≥17 years) and pediatric patients with nondisease conditions (eg, trauma, poisoning, or burns). Emergency physicians selected 592 representative cases for training and 50 for validation. Exposure A commercially available text generation transformer model was used as a core LLM, fine-tuned using the 592 training cases. Two distinct processing pipelines were implemented within the LLM assistant due to different input data: (1) for patients managed solely by emergency physicians, using the ED initial record and prescription list, and (2) for those requiring specialty consultations, using the ED initial record and consultation request form. Main Outcomes and Measures Quality of notes using 4C metrics (completeness, correctness, conciseness, and clinical utility) on a Likert scale ranging from 1 to 5 and time taken to complete the notes manually and with the LLM assistant. Results Of the 50 test cases, the mean (SD) patient age was 57.7 (23.1) years, and 28 patients (56%) were female. LLM-assisted notes achieved higher scores than manual notes in completeness (4.23 95% CI, 4.17-4.28 vs 4.03 95% CI, 3.96-4.09), correctness (4.38 95% CI, 4.33-4.42 vs 4.20 95% CI, 4.14-4.26), conciseness (4.23 95% CI, 4.18-4.28 vs 4.11 95% CI, 4.05-4.17), and clinical utility (4.17 95% CI, 4.11-4.23 vs 3.85 95% CI, 3.78-3.91) (all P lt; .001). When compared with LLM drafts, LLM-assisted notes excelled in conciseness (4.23 vs 3.98 95% CI, 3.91-4.04; P lt; .001) and maintained equivalent clinical utility (4.17 vs 4.16 95% CI, 4.11-4.21; P gt; .99), but scored lower in completeness (4.23 vs 4.34 95% CI, 4.29-4.39; P = .001) and correctness (4.38 vs 4.45 95% CI, 4.41-4.49; P lt; .001). The median documentation time per note dropped from 69.5 (95% CI, 65.5-78.0) seconds for manual notes to 32.0 (95% CI, 29.5-36.0) seconds for LLM-assisted notes ( P lt; .001). Conclusion In this comparative effectiveness study, use of an on-site LLM assistant was associated with reduced writing time for ED discharge notes compared with manual note-taking, without compromising documentation quality, representing a critical advancement in the use of artificial intelligence for clinical practice.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Song et al. (2025) studied this question.

synapsesocial.com/papers/68f9d6583f378872224927e5https://doi.org/10.1001/jamanetworkopen.2025.38427
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Evaluation Framework of Large Language Models in Medical Documentation: Development and Usability Study2024 · 48 citations
  2. 2Patient-Representing Population's Perceptions of GPT-Generated Versus Standard Emergency Department Discharge Instructions: Randomized Blind Survey Assessment2024 · 33 citations
  3. 3Improving the quality of discharge summaries through a direct feedback system2020 · 37 citations
  4. 4Physician Burnout in the Electronic Health Record Era: Are We Ignoring the Real Cause?2018 · 425 citations
  5. 5Adapted large language models can outperform medical experts in clinical text summarization2024 · 744 citations