Key result
ChatGPT-3.5 yields highly accurate pediatric kidney transplant responses but ~32% pose potential risks requiring human oversight.
Why the study?
Educating adolescents undergoing kidney transplantation requires extensive resources, and large language models like ChatGPT-3.5 may offer potential assistance for providing information to patients and caregivers.
Can ChatGPT-3.5 provide accurate, relevant, and safe information for adolescents and caregivers regarding pediatric kidney transplantation?
Can ChatGPT-3.5 provide accurate, relevant, and safe information for adolescents and caregivers regarding pediatric kidney transplantation?
ChatGPT-3.5 provides generally accurate and relevant information on pediatric kidney transplantation, but the presence of potentially risky outputs necessitates human oversight.
AI chatbots may aid transplant education; leaves open need for safety validation before clinical use.
BACKGROUND: Education and enhancing the knowledge of adolescents who will undergo kidney transplantation are among the primary objectives of their care. While there are specific interventions in place to achieve this, they require extensive resources. The rise of large language models like ChatGPT-3.5 offers potential assistance for providing information to patients. This study aimed to evaluate the accuracy, relevance, and safety of ChatGPT-3.5's responses to patient-centered questions about pediatric kidney transplantation. The objective was to assess whether ChatGPT-3.5 could be a supplementary educational tool for adolescents and their caregivers in a complex medical context. METHODS: A total of 37 questions about kidney transplantation were presented to ChatGPT-3.5, which was prompted to respond as a health professional would to a layperson. Five pediatric nephrologists independently evaluated the outputs for accuracy, relevance, comprehensiveness, understandability, readability, and safety. RESULTS: The mean accuracy, relevancy, and comprehensiveness scores for all outputs were 4.51, 4.56, and 4.55, respectively. Out of 37 outputs, four were rated as completely accurate, and seven were completely relevant and comprehensive. Only one output had an accuracy, relevancy, and comprehensiveness score below 4. Twelve outputs were considered potentially risky, but only three had a risk grade of moderate or higher. Outputs that were considered risky had an accuracy and relevancy below the average. CONCLUSION: Our findings suggest that ChatGPT could be a useful tool for adolescents or caregivers of individuals waiting for kidney transplantation. However, the presence of potentially risky outputs underscores the necessity for human oversight and validation.
No takes yet. Share an insight, caveat, or question.
Demirbaş et al. (2025) studied Pediatric kidney transplantation. ChatGPT-3.5 was evaluated on Accuracy, relevancy, and comprehensiveness of responses. ChatGPT-3.5 responses to pediatric kidney transplant questions showed high mean accuracy (4.51) and relevancy (4.56), but 12 of 37 outputs were potentially risky, requiring human oversight.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: