Results 21 to 30 of about 14,313,923 (129)
Deep Learning for Image Caption Generation for the Visually Impaired [PDF]
Individuals with challenges related to vision have to deal with the pervasiveness of complete or partial unemployment amongst their peer group. World Health Organization (WHO) had estimated that people with headcount surpassing 2 billion are going ...
Vajpeyee, Amit
core +3 more sources
Automatic Caption Generation for Aerial Images: A Survey [PDF]
Aerial images have attracted attention from researcher community since long time. Generating a caption for an aerial image describing its content in comprehensive way is less studied but important task as it has applications in agriculture, defence ...
Satone, Manisha P.; Matoshri College of Engineering and Research Centre, Nashik, India affiliated to Savitribai Phule Pune University +2 more
core +1 more source
Image Caption Generation via Unified Retrieval and Generation-Based Method
Image captioning is a multi-modal transduction task, translating the source image into the target language. Numerous dominant approaches primarily employed the generation-based or the retrieval-based method.
Shanshan Zhao +4 more
doaj +1 more source
Self-Learning for Few-Shot Remote Sensing Image Captioning
Large-scale caption-labeled remote sensing image samples are expensive to acquire, and the training samples available in practical application scenarios are generally limited.
Haonan Zhou +3 more
doaj +1 more source
Image Description Generator using Residual Neural Network and Long Short-Term Memory [PDF]
Human beings can describe scenarios and objects in a picture through vision easily whereas performing the same task with a computer is a complicated one.
Mahesh Kumar Morampudi +3 more
doaj +1 more source
Classification and object recognition in image processing has significantly improved computer vision tasks. The method is often used for visual problems, especially in picture classification utilizing the Convolutional Neural Network (CNN).
Rifqi Mulyawan +2 more
doaj +1 more source
3M: Multi-style image caption generation using Multi-modality features under Multi-UPDOWN model
In this paper, we build a multi-style generative model for stylish image captioning which uses multi-modality image features, ResNeXt features, and text features generated by DenseCap.
Chengxi Li, Brent Harrison
doaj +1 more source
A transformer-based Urdu image caption generation [PDF]
Image caption generation has emerged as a remarkable development that bridges the gap between Natural Language Processing (NLP) and Computer Vision (CV).
NAWAZ, Raheel +7 more
core +4 more sources
Vision Transformer and Language Model Based Radiology Report Generation
Recent advancements in transformers exploited computer vision problems which results in state-of-the-art models. Transformer-based models in various sequence prediction tasks such as language translation, sentiment classification, and caption generation ...
Mashood Mohammad Mohsan +5 more
doaj +1 more source
Sparse Adversarial Examples Attacking on Video Captioning Model [PDF]
Despite the fact that multi-modal deep learning such as image captioning model has been proved to be vulnerable to adversarial examples,the adversarial susceptibility in video caption generation is under-examined.There are two main reasons for this.On ...
QIU Jiangxing, TANG Xueming, WANG Tianmei, WANG Chen, CUI Yongquan, LUO Ting
doaj +1 more source

