Results 131 to 140 of about 43,272 (165)
Some of the next articles are maybe not open access.
Context based multimodal fusion
Proceedings of the 6th international conference on Multimodal interfaces, 2004We present a generic approach to multimodal fusion which we call context based multimodal integration. Key to this approach is that every multimodal input event is interpreted and enriched with respect to its local turn context. This local turn context comprises all previously recognized input events and the dialogue state that both belong to the same ...
openaire +1 more source
Deep Multimodal Fusion: A Hybrid Approach
International Journal of Computer Vision, 2017zbMATH Open Web Interface contents unavailable due to conflicting licenses.
Mohamed R. Amer +5 more
openaire +1 more source
Fusion engines for multimodal input
Proceedings of the 2009 international conference on Multimodal interfaces, 2009Fusion engines are fundamental components of multimodal inter-active systems, to interpret input streams whose meaning can vary according to the context, task, user and time. Other surveys have considered multimodal interactive systems; we focus more closely on the design, specification, construction and evaluation of fusion engines. We first introduce
Denis Lalanne +5 more
openaire +1 more source
Proceedings of the 8th international conference on Multimodal interfaces, 2006
This is a new hybrid fusion strategy based primarily on the implementation of two former and differentiated approaches to multimodal fusion [11] in multimodal dialogue systems. Both approaches, their predecessors and their respective advantages and disadvantages will be described in order to illustrate how the new strategy merges them into a more solid
Pilar Manchón Portillo +2 more
openaire +1 more source
This is a new hybrid fusion strategy based primarily on the implementation of two former and differentiated approaches to multimodal fusion [11] in multimodal dialogue systems. Both approaches, their predecessors and their respective advantages and disadvantages will be described in order to illustrate how the new strategy merges them into a more solid
Pilar Manchón Portillo +2 more
openaire +1 more source
Fuzzy Fusion in Multimodal Biometric Systems
2007Multimodal authentication systems represent an emerging trend for information security. These systems could replace conventional mono-modal biometric methods using two or more features for robust biometric authentication tasks. They employ unique combinations of measurable physical characteristics: fingerprint, facial features, iris of the eye, voice ...
Vincenzo Conti +4 more
openaire +3 more sources
Interpretable Multimodal Capsule Fusion
IEEE/ACM Transactions on Audio, Speech, and Language Processing, 2022Jianfeng Wu, Sijie Mai, Haifeng Hu 0001
openaire +1 more source
Hardware-accelerated multimodality volume fusion
SPIE Proceedings, 2005In this paper, we propose a novel technique of multimodality volume fusion using a graphics hardware. Our 3D texture based volume fusion algorithm consists of three steps: First, two volumes of different modalities are loaded into the texture memory in the GPU.
Helen Hong +3 more
openaire +1 more source
2010
This chapter gives an overview of multimodal information fusion from the machine-learning perspective. Humans interact with each other using different modalities of communication. These include speech, gestures, documents, etc. It is therefore natural that human-computer interaction (HCI) should facilitate the same multimodal form of communication.
Poh, N, Kittler, J
openaire +3 more sources
This chapter gives an overview of multimodal information fusion from the machine-learning perspective. Humans interact with each other using different modalities of communication. These include speech, gestures, documents, etc. It is therefore natural that human-computer interaction (HCI) should facilitate the same multimodal form of communication.
Poh, N, Kittler, J
openaire +3 more sources
Proceedings of the 2018 Workshop on Audio-Visual Scene Understanding for Immersive Multimedia, 2018
Two-hour movie or a short movie clip as its subset is intended to capture and present a meaningful (or significant) story in video to be recognized and understood by human audience. What if we substitute the task of human audience with that of an intelligent machine or robot capable of capturing and processing the semantic information in terms of audio
openaire +1 more source
Two-hour movie or a short movie clip as its subset is intended to capture and present a meaningful (or significant) story in video to be recognized and understood by human audience. What if we substitute the task of human audience with that of an intelligent machine or robot capable of capturing and processing the semantic information in terms of audio
openaire +1 more source

