Back to search

Article

OCR-Mediated Modality Dominance in Vision-Language Models: Implications for Radiology AI Trustworthiness

2026-02-24

Abstract excerpt

1. <h4>Background</h4> Vision-language models (VLMs) are increasingly proposed for radiologic decision support, yet the security implications of deploying general-domain, OCR-capable models in diagnostic workflows remain poorly characterized. When image-embedded text is not treated as untrusted input, the visual channel becomes vulnerable to adversarial manipulation through OCR-readable overlays. <h4>Methods</h4>...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
8bfbb01a-8cc6-5cc6-8257-a9b3cf95b628
DOI
10.64898/2026.02.22.26346828
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
OCR-Mediated Modality Dominance in Vision-Language Models: Implications for Radiology AI TrustworthinessDOI 10.64898/2026.02.22.26346828
Select a neighboring publication to make it the new centre.