Back to search

Article

Benchmarking language models for postgraduate medical education on the Italian national specialty examination

2026-06-22

Abstract excerpt

<title>Abstract</title> <p>Large language models (LLMs) are increasingly considered as educational tools for medical trainees, yet systematic evidence on their accuracy, robustness, and deployment feasibility remains limited. We evaluated 18 LLMs, specifically six proprietary, six open-source (1.7B–8B parameters), and six quantized variants of medically fine-tuned models, on the Italian national medical specialty...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
1ef7f4f1-d2ec-5f65-8192-3ac9f06f4755
DOI
10.21203/rs.3.rs-9722612/v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Benchmarking language models for postgraduate medical education on the Italian national specialty examinationDOI 10.21203/rs.3.rs-9722612/v1
Select a neighboring publication to make it the new centre.