Results 11 to 20 of about 11,822,367 (241)
Multi-Band Image Caption Generation Method Based on Feature Fusion [PDF]
This study proposes a multi-band detection image caption generation method based on feature fusion to address the common problem of poor performance in describing nighttime scenes, occluded target scenes, and captured blurred images in existing image ...
HE Shan, LIN Suzhen, WANG Yanbo, LI Dawei
doaj +3 more sources
Middle-Level Attribute-Based Language Retouching for Image Caption Generation
Image caption generation is attractive research which focuses on generating natural language sentences to describe the visual content of a given image. It is an interdisciplinary subject combining computer vision (CV) and natural language processing (NLP)
Zhibin Guan +4 more
doaj +2 more sources
Geo-Aware Image Caption Generation [PDF]
Standard image caption generation systems produce generic descriptions of images and do not utilize any contextual information or world knowledge. In particular, they are unable to generate captions that contain references to the geographic context of an
ILS LLI +6 more
core +2 more sources
Explicit Image Caption Editing
Given an image and a reference caption, the image caption editing task aims to correct the misalignment errors and generate a refined caption. However, all existing caption editing works are implicit models, ie, they directly produce the refined captions
Niu, Yulei +6 more
core +2 more sources
Image Captioning Using Motion-CNN with Object Detection
Automatic image captioning has many important applications, such as the depiction of visual contents for visually impaired people or the indexing of images on the internet.
Kiyohiko Iwamura +4 more
doaj +1 more source
VSAM-Based Visual Keyword Generation for Image Caption
Image caption is to understand and describe the visual content, which is expected to be applied in automatic news reporting in future. In recent years, there has been an increasing interest in an Encoder-Decoder framework for image caption: the encoder ...
Suya Zhang +3 more
doaj +1 more source
A core region captioning framework for automatic video understanding in story video contents
Due to the rapid increase in images and image data, research examining the visual analysis of such unstructured data has recently come to be actively conducted. One of the representative image caption models the DenseCap model extracts various regions in
Hyesun Suh +3 more
doaj +1 more source
Image-Caption Model Based on Fusion Feature
The encoder–decoder framework is the main frame of image captioning. The convolutional neural network (CNN) is usually used to extract grid-level features of the image, and the graph convolutional neural network (GCN) is used to extract the image’s ...
Yaogang Geng +3 more
doaj +1 more source
Arabic Captioning for Images of Clothing Using Deep Learning
Fashion is one of the many fields of application that image captioning is being used in. For e-commerce websites holding tens of thousands of images of clothing, automated item descriptions are quite desirable.
Rasha Saleh Al-Malki +1 more
doaj +1 more source
Learning from Children: Improving Image-Caption Pretraining via Curriculum [PDF]
Image-caption pretraining has been quite successfully used for downstream vision tasks like zero-shot image classification and object detection. However, image-caption pretraining is still a hard problem -- it requires multiple concepts (nouns) from ...
Zareian, Alireza +4 more
core +1 more source

