Back to search

Article

Evaluation of DeepSeek-R1 and ChatGPT on the Chinese National Medical Licensing Examination: A Multi-Year Comparative Study

2025-09-09

Abstract excerpt

<title>Abstract</title> <p> <bold>Background</bold> Large language models (LLMs) have demonstrated remarkable capabilities in natural language understanding and reasoning. However, their real-world applicability in high-stakes medical assessments remains underexplored, particularly in non-English contexts. This study aims to evaluate the performance of DeepSeek-R1 and ChatGPT on the Chinese National Medical Lic...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
7f45ce8c-f87a-5181-a234-090aa856a21a
DOI
10.21203/rs.3.rs-7380847/v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Evaluation of DeepSeek-R1 and ChatGPT on the Chinese National Medical Licensing Examination: A Multi-Year Comparative StudyDOI 10.21203/rs.3.rs-7380847/v1
Select a neighboring publication to make it the new centre.