Results 11 to 20 of about 11,822,367 (241)

Multi-Band Image Caption Generation Method Based on Feature Fusion [PDF]

open access: yesJisuanji gongcheng
This study proposes a multi-band detection image caption generation method based on feature fusion to address the common problem of poor performance in describing nighttime scenes, occluded target scenes, and captured blurred images in existing image ...
HE Shan, LIN Suzhen, WANG Yanbo, LI Dawei
doaj   +3 more sources

Middle-Level Attribute-Based Language Retouching for Image Caption Generation

open access: yesApplied Sciences, 2018
Image caption generation is attractive research which focuses on generating natural language sentences to describe the visual content of a given image. It is an interdisciplinary subject combining computer vision (CV) and natural language processing (NLP)
Zhibin Guan   +4 more
doaj   +2 more sources

Geo-Aware Image Caption Generation [PDF]

open access: yes, 2020
Standard image caption generation systems produce generic descriptions of images and do not utilize any contextual information or world knowledge. In particular, they are unable to generate captions that contain references to the geographic context of an
ILS LLI   +6 more
core   +2 more sources

Explicit Image Caption Editing

open access: yes, 2022
Given an image and a reference caption, the image caption editing task aims to correct the misalignment errors and generate a refined caption. However, all existing caption editing works are implicit models, ie, they directly produce the refined captions
Niu, Yulei   +6 more
core   +2 more sources

Image Captioning Using Motion-CNN with Object Detection

open access: yesSensors, 2021
Automatic image captioning has many important applications, such as the depiction of visual contents for visually impaired people or the indexing of images on the internet.
Kiyohiko Iwamura   +4 more
doaj   +1 more source

VSAM-Based Visual Keyword Generation for Image Caption

open access: yesIEEE Access, 2021
Image caption is to understand and describe the visual content, which is expected to be applied in automatic news reporting in future. In recent years, there has been an increasing interest in an Encoder-Decoder framework for image caption: the encoder ...
Suya Zhang   +3 more
doaj   +1 more source

A core region captioning framework for automatic video understanding in story video contents

open access: yesInternational Journal of Engineering Business Management, 2022
Due to the rapid increase in images and image data, research examining the visual analysis of such unstructured data has recently come to be actively conducted. One of the representative image caption models the DenseCap model extracts various regions in
Hyesun Suh   +3 more
doaj   +1 more source

Image-Caption Model Based on Fusion Feature

open access: yesApplied Sciences, 2022
The encoder–decoder framework is the main frame of image captioning. The convolutional neural network (CNN) is usually used to extract grid-level features of the image, and the graph convolutional neural network (GCN) is used to extract the image’s ...
Yaogang Geng   +3 more
doaj   +1 more source

Arabic Captioning for Images of Clothing Using Deep Learning

open access: yesSensors, 2023
Fashion is one of the many fields of application that image captioning is being used in. For e-commerce websites holding tens of thousands of images of clothing, automated item descriptions are quite desirable.
Rasha Saleh Al-Malki   +1 more
doaj   +1 more source

Learning from Children: Improving Image-Caption Pretraining via Curriculum [PDF]

open access: yes, 2023
Image-caption pretraining has been quite successfully used for downstream vision tasks like zero-shot image classification and object detection. However, image-caption pretraining is still a hard problem -- it requires multiple concepts (nouns) from ...
Zareian, Alireza   +4 more
core   +1 more source

Home - About - Disclaimer - Privacy