Results 21 to 30 of about 4,895,571 (286)
Knowledge-based Visual Question Answering:A Survey [PDF]
As an important presentation form of the completeness of artificial intelligence and the visual Turing test,visual question answering(VQA),coupled with its potential application value,has received extensive attention from computer vision and na-tural ...
WANG Ruiping, WU Shihong, ZHANG Meihang, WANG Xiaoping
doaj +1 more source
Generative Visual Question Answering
Multi-modal tasks involving vision and language in deep learning continue to rise in popularity and are leading to the development of newer models that can generalize beyond the extent of their training data. The current models lack temporal generalization which enables models to adapt to changes in future data.
Ethan Shen, Scotty Singh, Bhavesh Kumar
openaire +2 more sources
Question-Agnostic Attention for Visual Question Answering [PDF]
Visual Question Answering (VQA) models employ attention mechanisms to discover image locations that are most relevant for answering a specific question. For this purpose, several multimodal fusion strategies have been proposed, ranging from relatively simple operations (e.g., linear sum) to more complex ones (e.g., Block).
Moshiur R. Farazi +2 more
openaire +3 more sources
Pathological Visual Question Answering [PDF]
We develop datasets and methods to perform visual question answering on pathology images.
Xuehai He +6 more
openaire +2 more sources
Survey of Multimodal Medical Question Answering
Multimodal medical question answering (MMQA) is a vital area bridging healthcare and Artificial Intelligence (AI). This survey methodically examines the MMQA research published in recent years.
Hilmi Demirhan, Wlodek Zadrozny
doaj +1 more source
K-VQA: A visual question answering method
The types of questions answered by the visual question answering of images and texts are roughly divided into two types. The first type is the questions that can get the answers directly from the images, and the second type is the questions that need the
Hongbin GAO, Jinying MAO, Huiyong WANG
doaj +1 more source
Visual Question Answering: A Tutorial [PDF]
The task of visual question answering (VQA) is receiving increasing interest from researchers in both the computer vision and natural language processing fields. Tremendous advances have been seen in the field of computer vision due to the success of deep learning, in particular on low- and midlevel tasks, such as image segmentation or object ...
Damien Teney +2 more
openaire +2 more sources
Robust Explanations for Visual Question Answering [PDF]
WACV-2020 (Accepted)
Badri N. Patro +2 more
openaire +3 more sources
Visual Question Answering Method Based on Counterfactual Thinking [PDF]
Visual question answering(VQA) is a multi-modal task that combines computer vision and natural language proces-sing,which is extremely challenging.However,the current VQA model is often misled by the apparent correlation in the data,and the output of the
YUAN De-sen, LIU Xiu-jing, WU Qing-bo, LI Hong-liang, MENG Fan-man, NGAN King-ngi, XU Lin-feng
doaj +1 more source
An investigation of question translation for English-Chinese cross-language question answering [PDF]
Question-answering (QA) is a next-generation search technology which aims to provide answers to a user’s question from a collection of documents. Cross-Language QA (CLQA) extends this paradigm to answering questions from a collection in a different ...
Bin Wang +11 more
core +2 more sources

