Article
Performance of Advanced Large Language Models (GPT-4o, GPT-4, Gemini 1.5 Pro, Claude 3 Opus) on Japanese Medical Licensing Examination: A Comparative Study
2024-07-09
Abstract excerpt
<h4>Purpose</h4> This study aims to evaluate the accuracy of medical knowledge in the most advanced LLMs (GPT-4o, GPT-4, Gemini 1.5 Pro, and Claude 3 Opus) as of 2024. It is the first to evaluate these LLMs using a non-English medical licensing exam. The insights from this study will guide educators, policymakers, and technical experts in the effective use of AI in medical education and clinical diagnosis. <h4>Me...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- acfe3ab0-25df-59eb-a4e2-5cabf85cf9be
- DOI
- 10.1101/2024.07.09.24310129
