Back to search

Article

Towards Multimodal Retrieval-Augmented Generation for Medical Visual Question Answering

2025-10-28

Abstract excerpt

<title>Abstract</title> <p>Medical visual question answering (MedVQA) is a critical AI healthcare task that combines medical image analysis with natural language understanding to assist clinicians in decision-making. While medical vision-language models have shown promise in this domain, they struggle with factual inaccuracies and hallucinations. Retrieval-augmented generation (RAG) improves the factual accuracy...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
d2b3ba24-9f0e-555d-b8a2-57eafc75077e
DOI
10.21203/rs.3.rs-7752202/v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Towards Multimodal Retrieval-Augmented Generation for Medical Visual Question AnsweringDOI 10.21203/rs.3.rs-7752202/v1
Select a neighboring publication to make it the new centre.