Article
Comparative Performance of General-Purpose Large Language Models and a Specialized Clinical AI Tool on OKAP-Style Ophthalmology Questions
2026-07-08
Abstract excerpt
<title>Abstract</title> <p> <bold>Purpose</bold> To evaluate the performance, discrimination, and calibration of five contemporary large language models (LLMs), including OpenEvidence, on OKAP-style ophthalmology multiple-choice questions. <bold>Methods</bold> Five LLM-based systems were assessed: OpenEvidence, a specialized clinical AI platform, and four general-purpose LLMs: ChatGPT-5.4-Pro, Gemini-Pro-3, C...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 387d10d4-f413-5e2b-b48c-2045daee9ab7
- DOI
- 10.21203/rs.3.rs-10067407/v1
