論文 ·日本語 ·未確認
Evaluating Chat GPT-4o’s Comparative Performance over GPT-4 in Japanese Medical Licensing Examination and Its Clinical Partnership Potential
Masatoshi Miyamura ・ Goro Fujiki ・ Yumiko Kanzaki ・ Kosuke Tsuda ・ Hironaka Asano ・ Hideaki Morita ・ Masaaki Hoshiga
- 刊行年
- 2026-01-07
- 収録
- 『International Medical Education』 5(1) pp. 9-9
- 出版
- Multidisciplinary Digital Publishing Institute
- 言語
- 英語
- OpenAlex
- W7118665070
- DOI
- 10.3390/ime5010009
- ISSN
- 2813-141X
- URL
- https://www.mdpi.com/2813-141X/5/1/9/pdf?version=1767783012
要旨
Background: Recent advances in artificial intelligence (AI) have produced ChatGPT-4o, a multimodal large language model (LLM) capable of processing both text and image inputs. Although ChatGPT has demonstrated usefulness in medical examinations, few studies have evaluated its image analysis performance. Methods: This study compared GPT-4o and GPT-4 using public questions from the 116th–118th Japanese National Medical Licensing Examinations (JNMLE), each consisting of 400 questions. Both models answered in Japanese using simple prompts, including screenshots for image-based questions. Accuracy was analyzed across essential, general, and clinical questions, with statistical comparisons by chi-square tests. Results: GPT-4o consistently outperformed GPT-4, achieving passing scores in all three examinations. In the 118th JNMLE, GPT-4o scored 457 points versus 425 for GPT-4. GPT-4o demonstrated higher accuracy for image-based questions in the 117th and 116th exams, though the difference in the 118th was not significant. For text-based questions, GPT-4o showed superior medical knowledge, clinical reasoning, and ethical response behavior, notably avoiding prohibited options. Conclusion: Overall, GPT-4o exceeded GPT-4 in both text and image domains, suggesting strong potential as a diagnostic aid and educational resource. Its balanced performance across modalities highlights its promise for integration into future medical education and clinical decision support.
主題
この書誌の出所
- openalex— W7118665070(2026-08-14取得)
引用
Masatoshi Miyamura・Goro Fujiki・Yumiko Kanzaki・Kosuke Tsuda・Hironaka Asano・Hideaki Morita・Masaaki Hoshiga(2026-01-07) Evaluating Chat GPT-4o’s Comparative Performance over GPT-4 in Japanese Medical Licensing Examination and Its Clinical Partnership Potential 『International Medical Education』 5(1) pp. 9-9 Multidisciplinary Digital Publishing Institute