Key result
ChatGPT provides accurate heart attack advice, but ~41% of users follow recommendations without consulting physicians.
Why the study?
Does ChatGPT provide accurate and safe information regarding myocardial infarction, and how do users perceive and trust it?
Cross-Sectional (n=352)
Does ChatGPT provide accurate and safe information regarding myocardial infarction, and how do users perceive and trust it?
ChatGPT provides generally accurate and safe information about myocardial infarction but lacks completeness, highlighting its role as an adjunctive educational tool rather than a replacement for professional medical consultation.
Background: Artificial intelligence chatbots, particularly ChatGPT, have emerged as increasingly popular sources of health information for the general public. However, concerns persist regarding the accuracy, safety, and appropriateness of AI-generated medical advice, especially for life-threatening conditions such as myocardial infarction. Objectives: This study aimed to evaluate the accuracy, completeness, and safety of ChatGPT-generated responses and information to public questions about heart attacks, assess user perceptions and trust, and compare AI-generated content with expert-validated sources. Methods: A cross-sectional study was conducted between January and March 2026. A total of 73 commonly asked heart attack-related questions were submitted to ChatGPT (GPT-4), and responses were independently evaluated by two expert reviewers and one external reviewer using standardized rubric assessing accuracy, completeness, clarity, safety, and tone. Inter-rater reliability was assessed using intraclass correlation coefficients. In parallel, a survey of 352 participants evaluated user perceptions, trust, and behavioral use of ChatGPT for medical information. Statistical analyses included descriptive statistics, group comparisons, correlation analyses, and multivariable regression models to identify predictors of trust and satisfaction. Results: Expert reviewers assigned high ratings for accuracy (mean: 4.62 ± 0.62 and 4.48 ± 0.63) and safety (mean: 4.78 ± 0.51 and 4.05 ± 0.50), with substantial inter-rater agreement (ICC = 0.788). In contrast, the external reviewer documented considerably lower completeness ratings (2.59 ± 0.57), revealing deficiencies in information coverage. Remarkably, 40.6% of respondents indicated they had followed ChatGPT's medical recommendations without seeking professional consultation, while 17.6% reported depending on the chatbot during urgent medical situations. Levels of user trust and satisfaction were moderate, showing significant correlations with perceived accuracy, usage frequency, and healthcare-related educational background. Conclusions: Expert assessments confirmed that ChatGPT generally provides accurate and safe cardiovascular health information; nevertheless, considerable inter-reviewer variability-especially the markedly lower completeness scores from the external evaluator-reveals significant gaps in content depth. These results indicate that ChatGPT functions more appropriately as an adjunctive educational resource rather than a replacement for professional medical consultatio n, emphasizing the necessity for enhanced safety mechanisms, clearer user instructions, and standardized content protocols when addressing high-stakes health topics. ChatGPT should be considered an adjunctive resource rather than a substitute for urgent professional medical assessment in suspected heart attack cases.
No takes yet. Share an insight, caveat, or question.
Elsheikh et al. (2026) conducted a cross-sectional in Myocardial infarction (n=352). ChatGPT (GPT-4) vs. Expert-validated sources was evaluated on Accuracy, completeness, and safety of ChatGPT-generated responses, and user perceptions and trust. ChatGPT generated generally accurate and safe responses to heart attack questions, though 40.6% of surveyed users followed its medical recommendations without professional consultation.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: