Back to search

Article

Performance of o1 pro and GPT-4 in self-assessment questions for nephrology board renewal

2025-01-15

Abstract excerpt

<h4>ABSTRACT</h4> <h4>Background</h4> Large language models (LLMs) are increasingly evaluated in medical education and clinical decision support, but their performance in highly specialized fields, such as nephrology, is not well established. We compared two advanced LLMs, GPT-4 and the newly released o1 pro, on comprehensive nephrology board renewal examinations. <h4>Methods</h4> We administered 209 Japanese S...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
cb480ec5-b49f-5bb1-91e3-b57cf18fce83
DOI
10.1101/2025.01.14.25320525
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Performance of o1 pro and GPT-4 in self-assessment questions for nephrology board renewalDOI 10.1101/2025.01.14.25320525
Select a neighboring publication to make it the new centre.