PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
August 13, 20259 citations

Assessment of the Utility of Artificial Intelligence-Based Chatbots in Patient Education: A Systematic Review and Meta-Analysis.

View Full Paper
SESameh Hany EmileNHNir HoreshZGZoe Garoufalia

Key Points

  • Chatbots achieved a pooled appropriateness rate of 89.1% for patient education answers.
  • Responses from chatbots were accurate but had a high reading level, scoring an average of 13.1 on the Flesh-Kincaid Grade Level.
  • Chatbots showed 78.6%-95% concordance with published guidelines in colorectal surgery and urology.
  • Compared to Google Search, chatbots scored higher in patient education effectiveness, achieving 87% versus 78%.

Abstract

BackgroundChatbots and large language models, particularly ChatGPT, have led to an increasing number of studies on the potential for chatbots in patient education. In this systematic review, we aimed to provide a pooled assessment of the appropriateness and accuracy of chatbot responses in patient education across various medical disciplines.MethodsThis was a PRISMA-compliant systematic review and meta-analysis. PubMed and Scopus were searched from January-August 2023. Eligible studies that assessed the utility of chatbots in patient education were included. Primary outcomes were the appropriateness and quality of chatbot responses. Secondary outcomes included readability and concordance with published guidelines and Google searches. A random-effect proportional meta-analysis was used for pooling data.ResultsFollowing initial screening, 21 studies were included. The pooled rate of appropriateness of chatbot answers was 89.1% (95%CI: 84.9%-93.3%). ChatGPT was the most assessed chatbot. Responses, while accurate, were found to be at a college reading level as the weighted mean Flesh-Kincaid Grade Level was 13.1 (95%CI: 11.7-14.5) and the weighted mean Flesch Reading Ease Score was 38.6 (95%CI: 29- 48.2). Answers of chatbots to questions relevant to patient education had 78.6%-95% concordance with published guidelines in colorectal surgery and urology. Chatbots had higher patient education scores (87% vs 78%) than Google Search.ConclusionsChatbots provide largely accurate and appropriate answers for patient education. The advanced reading level of chatbot responses might be a limitation to their wide adoption as a source for patient education. However, they outperform traditional search engines and align well with professional guidelines, showcasing their potential in patient education.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Emile et al. (2025) studied this question.

synapsesocial.com/papers/689e03efd61984b91e13d45fhttps://doi.org/10.1177/00031348251367031
Ask AI
Helpful
Bookmark
Share
View Full Paper