PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 16, 2026Istanbul Medical Journal0 citationsOpen Access

Performance Evaluation of Large Language Models in Emergency Medicine Specialty Examination Questions: A Cross-Sectional Study

ŞKŞebnem Zeynep Eke KurtSBSuphi Bahadırlı

Key Points

  • This research aims to evaluate the performance of large language models (LLMs) in emergency medicine specialty examinations using the TUS format.
  • Cross-sectional study design comparing LLM performance to traditional examination formats.
  • Analyzed responses to TUS questions in the context of linguistic and curricular differences.
  • Focused on emergency medicine specialty examinations to assess contextual relevance.
  • LLMs demonstrated varying levels of performance when answering TUS questions compared to USMLE-style exams.
  • Performance metrics indicated strengths and weaknesses specific to the emergency medicine context, influencing assessment validity.

Abstract

Unlike prior studies primarily focused on USMLE-style examinations, this study evaluates LLM performance using the TUS, which reflects a different linguistic and curricular context.By directly comparing LLMs

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Kurt et al. (2026) studied this question.

synapsesocial.com/papers/6a0808afa487c87a6a40af33https://doi.org/10.4274/imj.galenos.2026.04742
Ask AI
Helpful
Bookmark
Share
View Full Paper