PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
August 20, 2025Australian Endodontic Journal0 citations

Assessment of the Accuracy of Modern Artificial Intelligence Chatbots in Responding to Endodontic Queries

View Full Paper
MÇMelis ÇakarAAAyşe Tuğba Eminsoy AvcıSDSalih Düzgün

Key Points

  • ChatGPT‐3.5 achieved the highest accuracy at 80%, indicating superior performance over other AI models.
  • The study involved 40 yes/no questions across 12 endodontic topics to evaluate chatbot responses against expert consensus.
  • Expert consensus was assessed using Cohen's kappa test, revealing weak agreement for ChatGPT and minimal for Gemini Flash.
  • Findings highlight the need for caution in using AI-generated responses for clinical decision-making in endodontics.

Abstract

ABSTRACT This study aims to compare the accuracy of modern AI chatbots, including Gemini 1.5 Flash, Gemini 1.5 Pro, ChatGPT‐3.5 and ChatGPT‐4, in responding to endodontic questions and supporting clinicians. Forty yes/no questions covering 12 endodontic topics were formulated by three experts. Each question was presented to the AI models on the same day, with a new chat session initiated for each. The agreement between chatbot responses and expert consensus was assessed using Cohen's kappa test ( p < 0.05). ChatGPT‐3.5 demonstrated the highest accuracy (80%), followed by ChatGPT‐4 (77.5%), Gemini 1.5 Pro (72.5%) and Gemini 1.5 Flash (60%). The agreement levels ranged from weak (ChatGPT models) to minimal (Gemini Flash). The findings indicate variability in chatbot performance, with ChatGPT models outperforming Gemini. However, reliance on AI‐generated responses for clinical decision‐making remains questionable. Future studies should incorporate more complex clinical scenarios and broader analytical approaches to enhance the assessment of AI chatbots in endodontics.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Çakar et al. (2025) studied this question.

synapsesocial.com/papers/68af56f4ad7bf08b1eadcdedhttps://doi.org/10.1111/aej.70012
Ask AI
Helpful
Bookmark
Share
View Full Paper