Back to search

Article

Comparative Performance of General-Purpose Large Language Models and a Specialized Clinical AI Tool on OKAP-Style Ophthalmology Questions

2026-07-08

Abstract excerpt

<title>Abstract</title> <p> <bold>Purpose</bold> To evaluate the performance, discrimination, and calibration of five contemporary large language models (LLMs), including OpenEvidence, on OKAP-style ophthalmology multiple-choice questions. <bold>Methods</bold> Five LLM-based systems were assessed: OpenEvidence, a specialized clinical AI platform, and four general-purpose LLMs: ChatGPT-5.4-Pro, Gemini-Pro-3, C...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
387d10d4-f413-5e2b-b48c-2045daee9ab7
DOI
10.21203/rs.3.rs-10067407/v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Comparative Performance of General-Purpose Large Language Models and a Specialized Clinical AI Tool on OKAP-Style Ophthalmology QuestionsDOI 10.21203/rs.3.rs-10067407/v1
Select a neighboring publication to make it the new centre.