Results 21 to 30 of about 13,237 (113)
Middle-Level Attribute-Based Language Retouching for Image Caption Generation
Image caption generation is attractive research which focuses on generating natural language sentences to describe the visual content of a given image. It is an interdisciplinary subject combining computer vision (CV) and natural language processing (NLP)
Zhibin Guan +4 more
doaj +1 more source
Classification and object recognition in image processing has significantly improved computer vision tasks. The method is often used for visual problems, especially in picture classification utilizing the Convolutional Neural Network (CNN).
Rifqi Mulyawan +2 more
doaj +1 more source
Cross-Lingual Image Caption Generation Based on Visual Attention Model
As an interesting and challenging problem, generating image caption automatically has attracted increasingly attention in natural language processing and computer vision communities.
Bin Wang +5 more
doaj +1 more source
3M: Multi-style image caption generation using Multi-modality features under Multi-UPDOWN model
In this paper, we build a multi-style generative model for stylish image captioning which uses multi-modality image features, ResNeXt features, and text features generated by DenseCap.
Chengxi Li, Brent Harrison
doaj +1 more source
Fine-Grained Remote Sensing Imagery Generation Driven by Expert Knowledge and Hierarchical Captions [PDF]
Current diffusion models struggle to achieve fine-grained remote sensing imagery (RSI) generation. This limitation fundamentally stems from their reliance on "flattened" text prompts, which overlook the inherent hierarchical structure of RSI.
J. Ren +12 more
doaj +1 more source
Multi-Band Image Caption Generation Method Based on Feature Fusion [PDF]
This study proposes a multi-band detection image caption generation method based on feature fusion to address the common problem of poor performance in describing nighttime scenes, occluded target scenes, and captured blurred images in existing image ...
HE Shan, LIN Suzhen, WANG Yanbo, LI Dawei
doaj +1 more source
Caption Generation Based on Emotions Using CSPDenseNet and BiLSTM with Self-Attention
Automatic image caption generation is an intricate task of describing an image in natural language by gaining insights present in an image. Featuring facial expressions in the conventional image captioning system brings out new prospects to generate ...
Kavi Priya S +5 more
doaj +1 more source
The video-based commonsense captioning task aims to add multiple commonsense descriptions to video captions to understand video content better. This paper aims to consider the importance of cross-modal mapping.
Haitao Xiong +4 more
doaj +1 more source
DEVELOPMENT OF IMAGE CAPTION GENERATION HYBRID MODEL
This study presents a hybrid model for image captioning using a VGG16 convolutional neural network (CNN) for feature extraction and a long short-term memory (LSTM) network for sequential text generation.
Aigul Mimenbayeva +2 more
doaj +1 more source
Content moderation assistance through image caption generation
The rapid growth in digital media creation has led to an increased challenge in content moderation. Manual and automated moderation are susceptible to risks associated with a slower response time and false positives arising from unpredictable user inputs
Liam Kearns
doaj +1 more source

