Abstract
This study evaluated the performance of ChatGPT with GPT-4 Omni (GPT-4o) on the 118th Japanese Medical Licensing Examination. The study focused on both text-only and image-based questions. The model demonstrated a high level of accuracy overall, with no significant difference in performance between text-only and image-based questions. Common errors included clinical judgment mistakes and prioritization issues, underscoring the need for further improvement in the integration of artificial intelligence into medical education and practice.
Author supplied keywords
Cite
CITATION STYLE
Miyazaki, Y., Hata, M., Omori, H., Hirashima, A., Nakagawa, Y., Eto, M., … Ikeda, M. (2024). Performance of ChatGPT-4o on the Japanese Medical Licensing Examination: Evalution of Accuracy in Text-Only and Image-Based Questions. JMIR Medical Education, 10. https://doi.org/10.2196/63129
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.