Authors
Loading...
Benchmarking study demonstrates variable accuracy and consistency across prompt formats on medical licensing questions, highlighting prompt sensitivity in AI education.
Carroll et al. (2026) studied this question.
Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context: