Article
Towards Multimodal Retrieval-Augmented Generation for Medical Visual Question Answering
2025-10-28
Abstract excerpt
<title>Abstract</title> <p>Medical visual question answering (MedVQA) is a critical AI healthcare task that combines medical image analysis with natural language understanding to assist clinicians in decision-making. While medical vision-language models have shown promise in this domain, they struggle with factual inaccuracies and hallucinations. Retrieval-augmented generation (RAG) improves the factual accuracy...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- d2b3ba24-9f0e-555d-b8a2-57eafc75077e
- DOI
- 10.21203/rs.3.rs-7752202/v1
