PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
October 2, 2025Frontiers in Digital Health68 citationsOpen Access

Large language models in real-world clinical workflows: a systematic review of applications and implementation

View Full Paper
YAYaara ArtsiVSVera SorinBGBenjamin S. Glicksberg

Key Points

  • LLMs showed notable improvements in operational efficiency and user satisfaction within clinical workflows.
  • Four peer-reviewed studies from 2024 to 2025 assessed LLMs, utilizing generative pre-trained transformers for various tasks.
  • A systematic review was conducted across several databases, emphasizing empirical implementations and limitations of LLMs.
  • Significant barriers to adoption include regulatory delays, performance variability, and the need for tailored implementation frameworks.

Abstract

Background Large language models (LLMs) offer promise for enhancing clinical care by automating documentation, supporting decision-making, and improving communication. However, their integration into real-world healthcare workflows remains limited and under characterized. This systematic review aims to evaluate the literature on real-world implementation of LLMs in clinical workflows, including their use cases, clinical settings, observed outcomes, and challenges. Methods We searched MEDLINE, Scopus, Web of Science, and Google Scholar for studies published between January 2015 and April 2025 that assessed LLMs in real-world clinical applications. Inclusion criteria were peer-reviewed, full-text studies in English reporting empirical implementation of LLMs in clinical settings. Study quality and risk of bias were assessed using the PROBAST tool. Results Four studies published between 2024 and 2025 met inclusion criteria. All used generative pre-trained transformers (GPTs). Reported applications included outpatient communication, mental health support, inbox message drafting, and clinical data extraction. LLM deployment was associated with improvements in operational efficiency, user satisfaction, and reduced workload. However, challenges included performance variability across data types, limitations in generalizability, regulatory delays, and lack of post-deployment monitoring. Conclusions Early evidence suggests that LLMs can enhance clinical workflows, but real-world adoption remains constrained by systemic, technical, and regulatory barriers. To support safe and scalable use, future efforts should prioritize standardized evaluation metrics, multi-site validation, human oversight, and implementation frameworks tailored to clinical settings. Systematic Review Registration https://www.crd.york.ac.uk/PROSPERO/recorddashboard , PROSPERO CRD420251030069.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Artsi et al. (2025) studied this question.

synapsesocial.com/papers/68de79595b556a9128e1a244https://doi.org/10.3389/fdgth.2025.1659134
Ask AI
Helpful
Bookmark
Share
View Full Paper