Back to search

Article

Benchmark Evaluation of Multi-Modal Large Language Models for Ophthalmic Diagnosis

2025-07-23

Abstract excerpt

<title>Abstract</title> <p>Multi-modal large language models (MLLMs) are increasingly demonstrating significant potential in medical applications, particularly in image-intensive fields such as ophthalmology. While state-of-the-art models like ChatGPT-4o and Qwen-VL 2.5 exhibit impressive performance in general-domain tasks, there remains a lack of real-world clinical benchmark datasets to rigorously evaluate the...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
910e2981-a8ff-58e7-9281-3b01e1c5edd2
DOI
10.21203/rs.3.rs-7186903/v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Benchmark Evaluation of Multi-Modal Large Language Models for Ophthalmic DiagnosisDOI 10.21203/rs.3.rs-7186903/v1
Select a neighboring publication to make it the new centre.