PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
December 1, 2023PLOS Digital Health79 citationsOpen Access

How does ChatGPT-4 preform on non-English national medical licensing examination? An evaluation in Chinese language

CFChangchang FangYWYuting WuWFWanying Fu

Key Points

Key points are not available for this paper at this time.

Abstract

ChatGPT, an artificial intelligence (AI) system powered by large-scale language models, has garnered significant interest in healthcare. Its performance dependent on the quality and quantity of training data available for a specific language, with the majority of it being in English. Therefore, its effectiveness in processing the Chinese language, which has fewer data available, warrants further investigation. This study aims to assess the of ChatGPT's ability in medical education and clinical decision-making within the Chinese context. We utilized a dataset from the Chinese National Medical Licensing Examination (NMLE) to assess ChatGPT-4's proficiency in medical knowledge in Chinese. Performance indicators, including score, accuracy, and concordance (confirmation of answers through explanation), were employed to evaluate ChatGPT's effectiveness in both original and encoded medical questions. Additionally, we translated the original Chinese questions into English to explore potential avenues for improvement. ChatGPT scored 442/600 for original questions in Chinese, surpassing the passing threshold of 360/600. However, ChatGPT demonstrated reduced accuracy in addressing open-ended questions, with an overall accuracy rate of 47.7%. Despite this, ChatGPT displayed commendable consistency, achieving a 75% concordance rate across all case analysis questions. Moreover, translating Chinese case analysis questions into English yielded only marginal improvements in ChatGPT's performance (p = 0.728). ChatGPT exhibits remarkable precision and reliability when handling the NMLE in Chinese. Translation of NMLE questions from Chinese to English does not yield an improvement in ChatGPT's performance.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Fang et al. (2023) studied this question.

synapsesocial.com/papers/6a0914cf0465d979db9d1ca3https://doi.org/10.1371/journal.pdig.0000397
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Performance of ChatGPT on Clinical Medicine Entrance Examination for Chinese Postgraduate in Chinese2023 · 5 citations
  2. 2Experiences, challenges, and prospects of National Medical Licensing Examination in China2022 · 44 citations
  3. 3Are ChatGPT's knowledge and interpretation ability comparable to those of medical students in Korea for taking a parasitology examination?: a descriptive study2023 · 201 citations
  4. 4Information and Artificial Intelligence2018 · 39 citations
  5. 5Performance of ChatGPT on USMLE: Potential for AI-assisted medical education using large language models2023 · 3,806 citations