Article
OCR-Mediated Modality Dominance in Vision-Language Models: Implications for Radiology AI Trustworthiness
2026-02-24
Abstract excerpt
1. <h4>Background</h4> Vision-language models (VLMs) are increasingly proposed for radiologic decision support, yet the security implications of deploying general-domain, OCR-capable models in diagnostic workflows remain poorly characterized. When image-embedded text is not treated as untrusted input, the visual channel becomes vulnerable to adversarial manipulation through OCR-readable overlays. <h4>Methods</h4>...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 8bfbb01a-8cc6-5cc6-8257-a9b3cf95b628
- DOI
- 10.64898/2026.02.22.26346828
