Article
Adaptive Semantic Fusion for Contextual Image Captioning
2025-01-21
Abstract excerpt
The automatic generation of textual descriptions from visual data is a fundamental yet challenging task that requires the seamless integration of image understanding and sophisticated language modeling. It involves not only identifying and interpreting complex visual elements but also effectively mapping them to coherent and contextually relevant textual representations. In this paper, we propose a novel framework...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- df0bb9ab-7c09-58ab-bf7f-06e2a91944dd
- DOI
- 10.20944/preprints202501.1491.v1
