Results 41 to 50 of about 4,895,571 (286)
Visual question answering with gated relation‐aware auxiliary
The great advances in computer vision and natural language processing make significant progress in visual question answering. In the visual question answering task, the visual representation is essential for understanding the image content.
Xiangjun Shao +2 more
doaj +1 more source
Co-Attention Network With Question Type for Visual Question Answering
Visual Question Answering (VQA) is a challenging multi-modal learning task since it requires an understanding of both visual and textual modalities simultaneously.
Chao Yang +4 more
doaj +1 more source
DOMAS: DATA ORIENTED MEDICAL VISUAL QUESTION ANSWERING USING SWIN TRANSFORMER
The Medical Visual Question Answering problem is a joined Computer Vision and Natural Language Processing task that aims to obtain answers in natural language to a question, posed in natural language as well, regarding an image.
Teodora-Alexandra TOADER
doaj +1 more source
Data and Knowledge Enhanced Medical Visual Question Answer Network [PDF]
Med-VQA aims to accurately answer clinical questions based on a given medical image,which is key in advancing clinical medical intelligence.Despite some progress in this field,challenges remain in extracting deep multimodal information from both images ...
YAN Yujing, HOU Xia, GUO Yuting, ZHANG Mingliang, SONG Wenfeng
doaj +1 more source
VISUAL QUESTIONING AND ANSWERING
In general, a VQA system is an algorithm that takes a picture and natural language query about the image as input and produces natural language response as output. The nature of a multi-discipline research problem necessitates this. This is a diagnostic dataset that assesses a variety of visual reason skills.
Lokesh R +3 more
openaire +1 more source
An Improved Attention for Visual Question Answering [PDF]
We consider the problem of Visual Question Answering (VQA). Given an image and a free-form, open-ended, question, expressed in natural language, the goal of VQA system is to provide accurate answer to this question with respect to the image. The task is challenging because it requires simultaneous and intricate understanding of both visual and textual ...
Tanzila Rahman +3 more
openaire +3 more sources
Intensional Question Answering using ILP: What does an answer mean? [PDF]
Cimiano P, Hartfiel H, Rudolph S. Intensional Question Answering using ILP: What does an answer mean? In: Kapetanios E, Sugumaran V, Spiliopoulou M, eds. Natural Language and Information Systems. Lecture Notes in Computer Science. Vol 5039.
Sugumaran, Vijayan +8 more
core +1 more source
Efficient question answering with question decomposition and multiple answer streams [PDF]
The German question answering (QA) system IRSAW (formerly: InSicht) participated in QA@CLEF for the fth time. IRSAW was introduced in 2007 by integrating the deep answer producer InSicht, several shallow answer producers, and a logical validator ...
Glöckner, Ingo +5 more
core +2 more sources
Multi-Question Learning for Visual Question Answering
Visual Question Answering (VQA) raises a great challenge for computer vision and natural language processing communities. Most of the existing approaches consider video-question pairs individually during training. However, we observe that there are usually multiple (either sequentially generated or not) questions for the target video in a VQA task, and
Chenyi Lei +6 more
openaire +3 more sources
An Open-Domain Question Answering System Using Annotated Web Feeds. [PDF]
Open Domain Question Answering Systems (ODQA) aim to answer all possible questions, regardless of topic and time. For this to be possible, most current ODQA systems depend on the World Wide Web through search engines, e.g.
Cheah, Yu-N +3 more
core +1 more source

