Results 121 to 130 of about 43,272 (165)
Some of the next articles are maybe not open access.
ACM Computing Surveys
Multimodal Artificial Intelligence (Multimodal AI), in general, involves various types of data (e.g., images, texts, or data collected from different sensors), feature engineering (e.g., extraction, combination/fusion), and decision-making (e.g., majority vote).
Chengcui Zhang +2 more
exaly +2 more sources
Multimodal Artificial Intelligence (Multimodal AI), in general, involves various types of data (e.g., images, texts, or data collected from different sensors), feature engineering (e.g., extraction, combination/fusion), and decision-making (e.g., majority vote).
Chengcui Zhang +2 more
exaly +2 more sources
2011
Multimodal interfaces offer its users the possibility of interacting with computers, in a transparent, natural way, by means of various modalities. Fusion engines are key components in multimodal systems, responsible for combining information from different sources and extract a semantic meaning from them.
Pedro Feiteira, Carlos Duarte
openaire +1 more source
Multimodal interfaces offer its users the possibility of interacting with computers, in a transparent, natural way, by means of various modalities. Fusion engines are key components in multimodal systems, responsible for combining information from different sources and extract a semantic meaning from them.
Pedro Feiteira, Carlos Duarte
openaire +1 more source
MobileFuse: Multimodal Image Fusion at the Edge
2023 26th International Conference on Information Fusion (FUSION), 2023The fusion of multiple images from different modalities is the process of generating a single output image that combines the useful information of all input images. Ideally, the information-rich content of each input image would be preserved, and the cognitive effort required by the user to extract this information should be smaller on the fused image ...
Perreault, H. +5 more
openaire +2 more sources
A system for multimodality image fusion
Proceedings of IEEE Symposium on Computer-Based Medical Systems (CBMS), 2002This paper describes a semi-automatic system for registering and visualizing CT (computer tomography) and MR (magnetic resonance) images. This system has been designed for images of the brain and augmented to include images of the spine. Registration requires identifying similar objects or structures in each image set.
Paul F. Hemler +4 more
openaire +1 more source
2009 IEEE International Conference on Multimedia and Expo, 2009
Our daily life involves multimodal signals (e.g., visual, audio, text, and signals from various sensors). Multimodal signals are highly correlated. For example, researchers have been using audio activities for video summarization of ball games and violence detection in movies. Multimodal signals are complementary to each other.
openaire +1 more source
Our daily life involves multimodal signals (e.g., visual, audio, text, and signals from various sensors). Multimodal signals are highly correlated. For example, researchers have been using audio activities for video summarization of ball games and violence detection in movies. Multimodal signals are complementary to each other.
openaire +1 more source
Multimodal image seamless fusion
Journal of Electronic Imaging, 2019An iterative joint bilateral filter is used to obtain a natural weight map. Images from different modalities are merged by a weighted-sum rule in the spatial domain. Saliency maps are determined by the gradient of the pairwise raw images. Comparing the pairwise values of saliency maps, a coarse weight map is attained to determine which pixel is ...
Kun Zhan, Lingwen Kong, Bo Liu, Ying He
openaire +1 more source
Fusion for Multimodal Biometric Identification
2005In this paper, we investigate fusion methods for multimodal identification using several unimodal identification results. One fingerprint identification system and two face identification systems are used as fusion sources. We discuss rank level and score level fusion methods.
Yongjin Lee +6 more
openaire +1 more source
Designing patterns for multimodal fusion
Proceedings of the 12th International Conference on Computer Systems and Technologies - CompSysTech '11, 2011This paper presents a design of an architecture that facilitates the work of a fusion engine. The logical combination or merging of input streams invoked by the fusion engine is based upon the definition of a set of patterns and its similarity with previously collected data from various modalities.
Ahmad Wehbi +2 more
openaire +1 more source
Multimodal Fusion of Visual Dialog
Proceedings of the 2020 2nd International Conference on Robotics, Intelligent Control and Artificial Intelligence, 2020Visual Dialog: aiming at holding a meaningful conversation with humans based on natural images, is a 'high-level' AI task of multimodal fusion. Since the challenge for visual dialog was proposed in 2017, multimodal fusion has been developed and made significant breakthroughs with the help of deep learning techniques.
Xiaofan Chen, Songyang Lao, Ting Duan
openaire +1 more source

