Back to search

Article

Compact Vision–Language Models Enable Efficient and Interpretable Automated OCT Analysis Through Layer Specific Multimodal Learning

2025-08-11

Abstract excerpt

Translating the intricate anatomical signatures of retinal disease from OCT B-scans into clear, accurate clinical narratives demands AI models that seamlessly fuse visual features with domain expertise. We curated a multimodal dataset of 40,000 OCT B-scans from public repositories and private clinical cohorts, each paired with expert validated summaries spanning six conditions: diabetic macular edema, diabetic ret...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
000e4199-1dc5-5f0b-9886-9f4584e4e331
DOI
10.1101/2025.08.07.669187
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Compact Vision–Language Models Enable Efficient and Interpretable Automated OCT Analysis Through Layer Specific Multimodal LearningDOI 10.1101/2025.08.07.669187
Select a neighboring publication to make it the new centre.