PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 12, 2026iScience0 citationsOpen Access

Comparison of the performance of large language models in answering patient questions related to cataract

View Full Paper
QHQing HeUniversity of Electronic Science and Technology of ChinaJSJiayi ShiHefei University of TechnologyXLXinyi LiuTiangong University

Key Points

Key points are not available for this paper at this time.

Abstract

This study evaluated the performance of four popular large-scale language models (ChatGPT o3-mini, Gemini 2.0 pro experimental, Deep Seek Thinking R1, and Kimi Thinking K1.5) in addressing frequently asked patient questions about cataracts and cataract surgery in Chinese. DeepSeek Thinking R1 performed comparably to Gemini 2.0 pro experimental in accuracy, while outperforming both ChatGPT o3-mini and Kimi Thinking K1.5. In terms of completeness and consistency, DeepSeek Thinking R1 showed superior performance over the other three LLMs. Regarding legibility and safety, DeepSeek Thinking R1, Gemini 2.0 pro experimental, and ChatGPT o3-mini exhibited comparable results, all performing better than Kimi Thinking K1.5. Deep Seek Thinking R1 demonstrated the strongest overall performance among the four LLMs in this comparative evaluation. The modern LLMs are promising tools for public education in ophthalmology while human oversight is still required.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

He et al. (2026) studied this question.

synapsesocial.com/papers/6a0889d0df3db87398109e9chttps://doi.org/10.1016/j.isci.2026.115002
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Ethical and regulatory challenges of large language models in medicine2024 · 326 citations
  2. 2Comparative analysis of the performance of the large language models DeepSeek-V3, DeepSeek-R1, open AI-O3 mini and open AI-O3 mini high in urology2025 · 24 citations
  3. 3Appropriateness and Comprehensiveness of Using ChatGPT for Perioperative Patient Education in Thoracic Surgery in Different Language Contexts: Survey Study2023 · 52 citations
  4. 4Readability and Appropriateness of Responses Generated by ChatGPT 3.5, ChatGPT 4.0, Gemini, and Microsoft Copilot for FAQs in Refractive Surgery2025 · 22 citations
  5. 5What the Canadian public (mis)understands about eyes and eye care2021 · 7 citations