Results 31 to 40 of about 14,313,923 (129)
Long-text caption generation for surgical image with a concept retrieval augmented large multimodal model [PDF]
Jiquan Liu +6 more
doaj +2 more sources
Image–text coherence and its implications for multimodal AI
Human communication often combines imagery and text into integrated presentations, especially online. In this paper, we show how image–text coherence relations can be used to model the pragmatics of image–text presentations in AI systems.
Malihe Alikhani +2 more
doaj +1 more source
Automatic caption generation with attention mechanisms aims at generating more descriptive captions containing coarser to finer semantic contents in the image.
Reshmi Sasibhooshan +2 more
doaj +1 more source
An Approach to Generate a Caption for an Image Collection Using Scene Graph Generation
Summarization is a challenging task that aims to generate a summary by grasping common information of a given set of information. Text summarization is a popular task of determining the topic or generating a textual summary of documents.
Itthisak Phueaksri +4 more
doaj +1 more source
A semantic event detection approach for soccer video based on perception concepts and finite state machines [PDF]
A significant application area for automated video analysis technology is the generation of personalized highlights of sports events. Sports games are always composed of a range of significant events.
Gareth J.F. Jones +9 more
core +2 more sources
Automatic creation of image descriptions, i.e. captioning of images, is an important topic in artificial intelligence (AI) that bridges the gap between computer vision (CV) and natural language processing (NLP).
Mukesh Kalla +3 more
doaj +1 more source
Efficient Explainable Metric for Semantic Communication of Images Using Image Captioning
As existing networks are projected to reach the theoretical limit for Shannon capacity, the concept of semantic communication is proposed in literature.
Asma O. Mahgoub, Elias Yaacoub
doaj +1 more source
Retrieval Topic Recurrent Memory Network for Remote Sensing Image Captioning
Remote sensing image (RSI) captioning aims to generate sentences to describe the content of RSIs. Generally, five sentences are used to describe the RSI in caption datasets.
Binqiang Wang +3 more
doaj +1 more source
Batik, a significant element of Indonesian cultural heritage, is renowned for its intricate patterns and profound philosophical meanings. While preserving traditional batik is crucial, the creation of modern patterns is equally encouraged to keep the art
Rahmatulloh Daffa Izzuddin Wahid +5 more
doaj +1 more source
Recent advances in deep learning and generative adversarial networks (GANs), in particular, has enabled interesting applications including photorealistic image generation, image translation, and automatic caption generation.
Sakib Shahriar, Noora Al Roken
doaj +1 more source

