Article
RefCap: Image Captioning with Referent Objects Attributes
2023-08-14
Abstract excerpt
In recent years, significant progress has been made in visual-linguistic multi-modality research, leading to advancements in visual comprehension and its applications in computer vision tasks. One fundamental task in visual-linguistic understanding is image captioning, which involves generating human-understandable textual descriptions given an input image. This paper introduces an end-to-end referring expression...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 1dc7367d-7433-5ac1-a809-f72f1ef36da1
- DOI
- 10.21203/rs.3.rs-3166028/v1
