Back to search

Article

Comparative accuracy of ChatGPT-o1, DeepSeek R1, and Gemini 2.0 in answering general primary care questions

2025-04-19

Abstract excerpt

<h4>Objectives</h4> To evaluate and compare the accuracy and reliability of large language models (LLMs) ChatGPT-o1, DeepSeek R1, and Gemini 2.0 in answering general primary care medical questions, assessing their reasoning approaches and potential applications in medical education and clinical decision-making. <h4>Design</h4> A cross-sectional study using an automated evaluation process where three large language...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
330e92a6-e0f1-5911-9997-66f14e72c506
DOI
10.1101/2025.04.15.25325518
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Comparative accuracy of ChatGPT-o1, DeepSeek R1, and Gemini 2.0 in answering general primary care questionsDOI 10.1101/2025.04.15.25325518
Select a neighboring publication to make it the new centre.