Article
Contextual Synergy through Explicit and Implicit Relations: A Unified Perspective for Image Description Generation
2025-10-08
Abstract excerpt
Automatic image description generation, commonly referred to as image captioning, has long been recognized as a highly demanding challenge within artificial intelligence, due to the necessity of bridging the perceptual gap between visual understanding and natural language expression. Conventional encoder-decoder pipelines typically convert salient image regions into textual sentences, yielding reasonable performan...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- d3349b12-70c9-5219-a1f2-3d167ca43c7d
- DOI
- 10.20944/preprints202510.0575.v1
