Back to search

Article

Evaluating the Performance of Large Language Models on a Neurology Board-Style Examination

2023-07-14

Abstract excerpt

<h4>Summary</h4> <h4>Background and Objectives</h4> Recent advancements in large language models (LLMs) such as GPT-3.5 and GPT-4 have shown impressive potential in a wide array of applications, including healthcare. While GPT-3.5 and GPT-4 showed heterogeneous results across specialized medical board examinations, the performance of these models in neurology board exams remains unexplored. <h4>Methods</h4> An exp...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
2206fb49-c4d6-5dcc-9e0c-ca2fbcc26ea4
DOI
10.1101/2023.07.13.23292598
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Evaluating the Performance of Large Language Models on a Neurology Board-Style ExaminationDOI 10.1101/2023.07.13.23292598
Select a neighboring publication to make it the new centre.