Results 31 to 40 of about 14,313,923 (129)

Long-text caption generation for surgical image with a concept retrieval augmented large multimodal model [PDF]

open access: yesPLoS ONE
Jiquan Liu   +6 more
doaj   +2 more sources

Image–text coherence and its implications for multimodal AI

open access: yesFrontiers in Artificial Intelligence, 2023
Human communication often combines imagery and text into integrated presentations, especially online. In this paper, we show how image–text coherence relations can be used to model the pragmatics of image–text presentations in AI systems.
Malihe Alikhani   +2 more
doaj   +1 more source

Image caption generation using Visual Attention Prediction and Contextual Spatial Relation Extraction

open access: yesJournal of Big Data, 2023
Automatic caption generation with attention mechanisms aims at generating more descriptive captions containing coarser to finer semantic contents in the image.
Reshmi Sasibhooshan   +2 more
doaj   +1 more source

An Approach to Generate a Caption for an Image Collection Using Scene Graph Generation

open access: yesIEEE Access, 2023
Summarization is a challenging task that aims to generate a summary by grasping common information of a given set of information. Text summarization is a popular task of determining the topic or generating a textual summary of documents.
Itthisak Phueaksri   +4 more
doaj   +1 more source

A semantic event detection approach for soccer video based on perception concepts and finite state machines [PDF]

open access: yes, 2007
A significant application area for automated video analysis technology is the generation of personalized highlights of sports events. Sports games are always composed of a range of significant events.
Gareth J.F. Jones   +9 more
core   +2 more sources

High-level and Low-level Feature Set for Image Caption Generation with Optimized Convolutional Neural Network

open access: yesJournal of Telecommunications and Information Technology, 2022
Automatic creation of image descriptions, i.e. captioning of images, is an important topic in artificial intelligence (AI) that bridges the gap between computer vision (CV) and natural language processing (NLP).
Mukesh Kalla   +3 more
doaj   +1 more source

Efficient Explainable Metric for Semantic Communication of Images Using Image Captioning

open access: yesIEEE Access
As existing networks are projected to reach the theoretical limit for Shannon capacity, the concept of semantic communication is proposed in literature.
Asma O. Mahgoub, Elias Yaacoub
doaj   +1 more source

Retrieval Topic Recurrent Memory Network for Remote Sensing Image Captioning

open access: yesIEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing, 2020
Remote sensing image (RSI) captioning aims to generate sentences to describe the content of RSIs. Generally, five sentences are used to describe the RSI in caption datasets.
Binqiang Wang   +3 more
doaj   +1 more source

Prompt Conditioned Batik Pattern Generation Using LoRA Weighted Diffusion Model With Classifier-Free Guidance

open access: yesIEEE Access
Batik, a significant element of Indonesian cultural heritage, is renowned for its intricate patterns and profound philosophical meanings. While preserving traditional batik is crucial, the creation of modern patterns is equally encouraged to keep the art
Rahmatulloh Daffa Izzuddin Wahid   +5 more
doaj   +1 more source

How can generative adversarial networks impact computer generated art? Insights from poetry to melody conversion

open access: yesInternational Journal of Information Management Data Insights, 2022
Recent advances in deep learning and generative adversarial networks (GANs), in particular, has enabled interesting applications including photorealistic image generation, image translation, and automatic caption generation.
Sakib Shahriar, Noora Al Roken
doaj   +1 more source

Home - About - Disclaimer - Privacy