AbstractIntroduction Traumatic dental injuries (TDIs) require prompt and accurate guidance, yet little is known about how the prompting influences the quality of artificial intelligence (AI) chatbot responses in these situations Methods Four dental trauma scenarios were developed by expert endodontists. For each scenario, a series of questions was posed to four AI chatbots (Claude Sonnet 3.5, Microsoft Copilot, GPT-4, Gemini Pro 2.5) using two approaches: unprompted layperson phrasing (n = 10) and endodontist's prompts referencing International Association of Dental Traumatology guidelines (n = 10). Responses were independently evaluated by two raters for validity, completeness, and relevance using a five-point ordinal scale. Results Scores (N=1920) of responses to endodontist questions were associated with significantly higher odds of receiving superior ratings across all three domains: validity (odds ratio OR = 1.82; 95% CI: 1.35–2.38; p Conclusions The quality of AI chatbot guidance in TDIs is significantly associated with how questions are asked. While clinically structured prompts yield more reliable responses, most patients facing dental trauma are unlikely to formulate questions in this way. This gap highlights an important limitation of current AI chatbot applications in TDIs and underscores the need for caution when relying on these tools without professional input.
Ourang et al. (Fri,) studied this question.