Results 1 to 10 of about 2,081,069 (199)
Robust Audio–Visual Speaker Localization in Noisy Aircraft Cabins for Inflight Medical Assistance [PDF]
Active Speaker Localization (ASL) involves identifying both who is speaking and where they are speaking from within audiovisual content. This capability is crucial in constrained and acoustically challenging environments, such as aircraft cabins during ...
Qiwu Qin, Yian Zhu
doaj +3 more sources
Existing noise-robust and reverberant-robust localization algorithms fail to localize the target speaker when interfering speakers are present. In this paper, we address the problem of localizing only the target speaker in multi-speaker scenarios and ...
Wenju Liu, Guanjun Li
exaly +3 more sources
Common sound source localization algorithms focus on localizing all the active sources in the environment. While the source identities are generally unknown, retrieving the location of a speaker of interest requires extra effort. This paper addresses the
Ziteng Wang, Junfeng Li, Yonghong Yan
doaj +4 more sources
Keyword Based Speaker Localization: Localizing a Target Speaker in a Multi-speaker Environment [PDF]
Speaker localization is a hard task, especially in adverse environmental conditions involving reverberation and noise. In this work we introduce the new task of localizing the speaker who uttered a given keyword, e.g., the wake-up word of a distant-microphone voice command system, in the presence of overlapping speech.
Sivasankaran, Sunit +2 more
openaire +3 more sources
Joint speaker localization and array calibration using expectation-maximization
Ad hoc acoustic networks comprising multiple nodes, each of which consists of several microphones, are addressed. From the ad hoc nature of the node constellation, microphone positions are unknown.
Yuval Dorfan +2 more
doaj +2 more sources
Three-Dimensional Speaker Localization: Audio-Refined Visual Scaling Factor Estimation
Neither a monocular RGB camera nor a small-size microphone array is capable of accurate three-dimensional (3D) speaker localization. By taking advantage of accurate visual object detection, and audio-visual complementary sensor fusion, we formulate the ...
Qi Liu, Haizhou Li, Xinyuan Qian
exaly +2 more sources
Weighted Frequency Smoothing for Enhanced Speaker Localization
The coherent signal subspace method may be used in order to apply subspace localization methods (e.g. MUSIC) to coherent sources. This method involves a focusing process followed by frequency smoothing, which is intended to decorrelate source signals ...
Hanan Beit-On, Tom Shlomo, Boaz Rafaely
openaire +2 more sources
Speaker identification and localization using shuffled MFCC features and deep learning
The use of machine learning in automatic speaker identification and localization systems has recently seen significant advances. However, this progress comes at the cost of using complex models, computations, and increasing the number of microphone ...
Anke Schmeink
exaly +2 more sources
Recent research advances in deep neural network (DNN)-based beamformers have shown great promise for speech enhancement under adverse acoustic conditions.
Hsinyu Chang +3 more
doaj +3 more sources
Audio Visual Speaker Localization from EgoCentric Views [PDF]
The use of audio and visual modality for speaker localization has been well studied in the literature by exploiting their complementary characteristics. However, most previous works employ the setting of static sensors mounted at fixed positions.
Jinzheng Zhao +3 more
semanticscholar +1 more source

