Article
Benchmark Evaluation of Multi-Modal Large Language Models for Ophthalmic Diagnosis
2025-07-23
Abstract excerpt
<title>Abstract</title> <p>Multi-modal large language models (MLLMs) are increasingly demonstrating significant potential in medical applications, particularly in image-intensive fields such as ophthalmology. While state-of-the-art models like ChatGPT-4o and Qwen-VL 2.5 exhibit impressive performance in general-domain tasks, there remains a lack of real-world clinical benchmark datasets to rigorously evaluate the...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 910e2981-a8ff-58e7-9281-3b01e1c5edd2
- DOI
- 10.21203/rs.3.rs-7186903/v1
