Results 21 to 30 of about 4,895,571 (286)

Knowledge-based Visual Question Answering:A Survey [PDF]

open access: yesJisuanji kexue, 2023
As an important presentation form of the completeness of artificial intelligence and the visual Turing test,visual question answering(VQA),coupled with its potential application value,has received extensive attention from computer vision and na-tural ...
WANG Ruiping, WU Shihong, ZHANG Meihang, WANG Xiaoping
doaj   +1 more source

Generative Visual Question Answering

open access: yesCoRR, 2023
Multi-modal tasks involving vision and language in deep learning continue to rise in popularity and are leading to the development of newer models that can generalize beyond the extent of their training data. The current models lack temporal generalization which enables models to adapt to changes in future data.
Ethan Shen, Scotty Singh, Bhavesh Kumar
openaire   +2 more sources

Question-Agnostic Attention for Visual Question Answering [PDF]

open access: yes2020 25th International Conference on Pattern Recognition (ICPR), 2021
Visual Question Answering (VQA) models employ attention mechanisms to discover image locations that are most relevant for answering a specific question. For this purpose, several multimodal fusion strategies have been proposed, ranging from relatively simple operations (e.g., linear sum) to more complex ones (e.g., Block).
Moshiur R. Farazi   +2 more
openaire   +3 more sources

Pathological Visual Question Answering [PDF]

open access: yes, 2020
We develop datasets and methods to perform visual question answering on pathology images.
Xuehai He   +6 more
openaire   +2 more sources

Survey of Multimodal Medical Question Answering

open access: yesBioMedInformatics, 2023
Multimodal medical question answering (MMQA) is a vital area bridging healthcare and Artificial Intelligence (AI). This survey methodically examines the MMQA research published in recent years.
Hilmi Demirhan, Wlodek Zadrozny
doaj   +1 more source

K-VQA: A visual question answering method

open access: yesJournal of Hebei University of Science and Technology, 2020
The types of questions answered by the visual question answering of images and texts are roughly divided into two types. The first type is the questions that can get the answers directly from the images, and the second type is the questions that need the
Hongbin GAO, Jinying MAO, Huiyong WANG
doaj   +1 more source

Visual Question Answering: A Tutorial [PDF]

open access: yesIEEE Signal Processing Magazine, 2017
The task of visual question answering (VQA) is receiving increasing interest from researchers in both the computer vision and natural language processing fields. Tremendous advances have been seen in the field of computer vision due to the success of deep learning, in particular on low- and midlevel tasks, such as image segmentation or object ...
Damien Teney   +2 more
openaire   +2 more sources

Robust Explanations for Visual Question Answering [PDF]

open access: yes2020 IEEE Winter Conference on Applications of Computer Vision (WACV), 2020
WACV-2020 (Accepted)
Badri N. Patro   +2 more
openaire   +3 more sources

Visual Question Answering Method Based on Counterfactual Thinking [PDF]

open access: yesJisuanji kexue, 2022
Visual question answering(VQA) is a multi-modal task that combines computer vision and natural language proces-sing,which is extremely challenging.However,the current VQA model is often misled by the apparent correlation in the data,and the output of the
YUAN De-sen, LIU Xiu-jing, WU Qing-bo, LI Hong-liang, MENG Fan-man, NGAN King-ngi, XU Lin-feng
doaj   +1 more source

An investigation of question translation for English-Chinese cross-language question answering [PDF]

open access: yes, 2007
Question-answering (QA) is a next-generation search technology which aims to provide answers to a user’s question from a collection of documents. Cross-Language QA (CLQA) extends this paradigm to answering questions from a collection in a different ...
Bin Wang   +11 more
core   +2 more sources

Home - About - Disclaimer - Privacy