本文へ移動

論文 ·日本語 ·未確認

Evaluating Chat GPT-4o’s Comparative Performance over GPT-4 in Japanese Medical Licensing Examination and Its Clinical Partnership Potential

Masatoshi Miyamura Goro Fujiki Yumiko Kanzaki Kosuke Tsuda Hironaka Asano Hideaki Morita Masaaki Hoshiga

刊行年
2026-01-07
収録
『International Medical Education』 5(1) pp. 9-9
出版
Multidisciplinary Digital Publishing Institute
言語
英語
OpenAlex
W7118665070
DOI
10.3390/ime5010009
ISSN
2813-141X
URL
https://www.mdpi.com/2813-141X/5/1/9/pdf?version=1767783012

要旨

Background: Recent advances in artificial intelligence (AI) have produced ChatGPT-4o, a multimodal large language model (LLM) capable of processing both text and image inputs. Although ChatGPT has demonstrated usefulness in medical examinations, few studies have evaluated its image analysis performance. Methods: This study compared GPT-4o and GPT-4 using public questions from the 116th–118th Japanese National Medical Licensing Examinations (JNMLE), each consisting of 400 questions. Both models answered in Japanese using simple prompts, including screenshots for image-based questions. Accuracy was analyzed across essential, general, and clinical questions, with statistical comparisons by chi-square tests. Results: GPT-4o consistently outperformed GPT-4, achieving passing scores in all three examinations. In the 118th JNMLE, GPT-4o scored 457 points versus 425 for GPT-4. GPT-4o demonstrated higher accuracy for image-based questions in the 117th and 116th exams, though the difference in the 118th was not significant. For text-based questions, GPT-4o showed superior medical knowledge, clinical reasoning, and ethical response behavior, notably avoiding prohibited options. Conclusion: Overall, GPT-4o exceeded GPT-4 in both text and image domains, suggesting strong potential as a diagnostic aid and educational resource. Its balanced performance across modalities highlights its promise for integration into future medical education and clinical decision support.

主題

この書誌の出所

  • openalex— W7118665070(2026-08-14取得)

引用

Masatoshi Miyamura・Goro Fujiki・Yumiko Kanzaki・Kosuke Tsuda・Hironaka Asano・Hideaki Morita・Masaaki Hoshiga(2026-01-07) Evaluating Chat GPT-4o’s Comparative Performance over GPT-4 in Japanese Medical Licensing Examination and Its Clinical Partnership Potential 『International Medical Education』 5(1) pp. 9-9 Multidisciplinary Digital Publishing Institute

Miyamura2026EvaluatingChatGPT
書誌 67,320件 語別索引 17,251件 資源 113件 研究者 303名 JSON