Results 21 to 30 of about 11,822,367 (241)

Localization and recognition of the scoreboard in sports video based on SIFT point matching [PDF]

open access: yes, 2011
In broadcast sports video, the scoreboard is attached at a fixed location in the video and generally the scoreboard always exists in all video frames in order to help viewers to understand the match’s progression quickly.
Guo, Jinlin   +4 more
core   +2 more sources

Image Caption Combined with GAN Training Method

open access: yes, 2020
Part 7: Computer Vision and Image UnderstandingInternational audienceIn today’s world where the number of images is huge and people cannot quickly retrieve the information they need, we urgently need a simpler and more human-friendly way of understanding
Shi, Zhongzhi   +3 more
core   +1 more source

Multifaceted Feature Coding Image Caption Generation Algorithm Based on Transformer [PDF]

open access: yesJisuanji gongcheng, 2023
Object features extracted by object detection algorithms play an increasingly critical role in the generation of image captions.However, only using the features of object detection as the input of an image caption task can lead to the loss of other ...
HENG Hongjun, FAN Yuchen, WANG Jialiang
doaj   +1 more source

Self-Learning for Few-Shot Remote Sensing Image Captioning

open access: yesRemote Sensing, 2022
Large-scale caption-labeled remote sensing image samples are expensive to acquire, and the training samples available in practical application scenarios are generally limited.
Haonan Zhou   +3 more
doaj   +1 more source

Cross-Lingual Image Caption Generation Based on Visual Attention Model

open access: yesIEEE Access, 2020
As an interesting and challenging problem, generating image caption automatically has attracted increasingly attention in natural language processing and computer vision communities.
Bin Wang   +5 more
doaj   +1 more source

Reconsidering Tourism Destination Images by Exploring Similarities between Travelogue Texts and Photographs

open access: yesISPRS International Journal of Geo-Information, 2022
With the rise of user-generated content (UGC) and deep learning technology, more and more researchers construct and measure the tourism destination image (TDI) through online travelogues. However, due to the impact of COVID-19 prevention and control, the
Xin Zhang   +3 more
doaj   +1 more source

Deconfounded fashion image captioning with transformer and multimodal retrieval

open access: yesVirtual Reality & Intelligent Hardware
Background: The annotation of fashion images is a significantly important task in the fashion industry as well as social media and e-commerce. However, owing to the complexity and diversity of fashion images, this task entails multiple challenges ...
Tao Peng   +4 more
doaj   +1 more source

Multilayer Dense Attention Model for Image Caption

open access: yesIEEE Access, 2019
The image caption is a technology that enables us to understand the contents and generate descriptive text, of images using machines. With the development of deep learning, means of using it to understand image content and generate descriptive text has ...
Ke Wang   +4 more
doaj   +1 more source

Generating Images from Caption and Vice Versa via CLIP-Guided Generative Latent Space Search [PDF]

open access: yes, 2021
In this research work we present CLIP-GLaSS, a novel zero-shot framework to generate an image (or a caption) corresponding to a given caption (or image).
Federico Galatolo   +5 more
core   +1 more source

VSRI:Visual Semantic Relational Interactor for Image Caption [PDF]

open access: yesJisuanji kexue
Image captioning is one of the key objectives of multimodal image understanding.This paper aims to generate detail-rich and accurate image caption.Currently,mainstream image captioning methods focus on the interrelationships between regions,but ignore ...
LIU Jian, YAO Renyuan, GAO Nan, LIANG Ronghua, CHEN Peng
doaj   +1 more source

Home - About - Disclaimer - Privacy