ChatGPT-3.5 provided comprehensive and correct responses to 78.2% of patient-generated queries regarding post-operative care following benign prostatic hyperplasia surgery.
Cross-Sectional (n=216)
Does ChatGPT-3.5 provide accurate and valid responses to patient-generated queries following BPH surgery?
ChatGPT-3.5 shows potential in answering post-operative BPH queries with 78.2% accuracy, though the risk of incomplete or misleading information necessitates further refinement.
The rapid advancement of artificial intelligence, particularly large language models like ChatGPT-3.5, presents promising applications in healthcare. This study evaluates ChatGPT-3.5's validity in responding to post-operative patient inquiries following surgery for benign prostatic hyperplasia (BPH). Common patient-generated questions were sourced from discharge instructions, online forums, and social media, covering various BPH surgical modalities. ChatGPT-3.5 responses were assessed by two senior urology residents using pre-defined criteria, with discrepancies resolved by a third reviewer. A total of 496 responses were reviewed, with 280 excluded. Among the 216 graded responses, 78.2% were comprehensive and correct, 9.3% were incomplete or partially correct, 10.2% contained a mix of accurate and inaccurate information, and 2.3% were entirely incorrect. Newer procedures (Aquablation, Rezum, iTIND) had a higher percentage of correct answers compared to traditional techniques (TURP, simple prostatectomy). The most common errors involved missing context or incorrect details (36.6%). These findings suggest that ChatGPT-3.5 has potential in providing accurate post-operative guidance for BPH patients. However, concerns regarding incomplete and misleading responses highlight the need for further refinement to improve AI-generated medical advice and ensure patient safety. Future research should focus on enhancing AI reliability in clinical applications.
Najdi et al. (Tue,) conducted a cross-sectional in Benign prostatic hyperplasia (BPH) (n=216). ChatGPT-3.5 was evaluated on Proportion of comprehensive and correct responses (Score 1). ChatGPT-3.5 provided comprehensive and correct responses to 78.2% of patient-generated queries regarding post-operative care following benign prostatic hyperplasia surgery.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: