Results 1 to 10 of about 2,081,069 (199)

Robust Audio–Visual Speaker Localization in Noisy Aircraft Cabins for Inflight Medical Assistance [PDF]

open access: yesSensors
Active Speaker Localization (ASL) involves identifying both who is speaking and where they are speaking from within audiovisual content. This capability is crucial in constrained and acoustically challenging environments, such as aircraft cabins during ...
Qiwu Qin, Yian Zhu
doaj   +3 more sources

GCC-Speaker: Target Speaker Localization with Optimal Speaker-Dependent Weighting in Multi-Speaker Scenarios

open access: yesICASSP 2023 - 2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), 2023
Existing noise-robust and reverberant-robust localization algorithms fail to localize the target speaker when interfering speakers are present. In this paper, we address the problem of localizing only the target speaker in multi-speaker scenarios and ...
Wenju Liu, Guanjun Li
exaly   +3 more sources

Target Speaker Localization Based on the Complex Watson Mixture Model and Time-Frequency Selection Neural Network

open access: yesApplied Sciences, 2018
Common sound source localization algorithms focus on localizing all the active sources in the environment. While the source identities are generally unknown, retrieving the location of a speaker of interest requires extra effort. This paper addresses the
Ziteng Wang, Junfeng Li, Yonghong Yan
doaj   +4 more sources

Keyword Based Speaker Localization: Localizing a Target Speaker in a Multi-speaker Environment [PDF]

open access: yesInterspeech 2018, 2018
Speaker localization is a hard task, especially in adverse environmental conditions involving reverberation and noise. In this work we introduce the new task of localizing the speaker who uttered a given keyword, e.g., the wake-up word of a distant-microphone voice command system, in the presence of overlapping speech.
Sivasankaran, Sunit   +2 more
openaire   +3 more sources

Joint speaker localization and array calibration using expectation-maximization

open access: yesEURASIP Journal on Audio, Speech, and Music Processing, 2020
Ad hoc acoustic networks comprising multiple nodes, each of which consists of several microphones, are addressed. From the ad hoc nature of the node constellation, microphone positions are unknown.
Yuval Dorfan   +2 more
doaj   +2 more sources

Three-Dimensional Speaker Localization: Audio-Refined Visual Scaling Factor Estimation

open access: yesIEEE Signal Processing Letters, 2021
Neither a monocular RGB camera nor a small-size microphone array is capable of accurate three-dimensional (3D) speaker localization. By taking advantage of accurate visual object detection, and audio-visual complementary sensor fusion, we formulate the ...
Qi Liu, Haizhou Li, Xinyuan Qian
exaly   +2 more sources

Weighted Frequency Smoothing for Enhanced Speaker Localization

open access: yesIEEE/ACM Transactions on Audio, Speech, and Language Processing, 2023
The coherent signal subspace method may be used in order to apply subspace localization methods (e.g. MUSIC) to coherent sources. This method involves a focusing process followed by frequency smoothing, which is intended to decorrelate source signals ...
Hanan Beit-On, Tom Shlomo, Boaz Rafaely
openaire   +2 more sources

Speaker identification and localization using shuffled MFCC features and deep learning

open access: yesInternational Journal of Speech Technology, 2023
The use of machine learning in automatic speaker identification and localization systems has recently seen significant advances. However, this progress comes at the cost of using complex models, computations, and increasing the number of microphone ...
Anke Schmeink
exaly   +2 more sources

Deep beamforming for speech enhancement and speaker localization with an array response-aware loss function

open access: yesFrontiers in Signal Processing
Recent research advances in deep neural network (DNN)-based beamformers have shown great promise for speech enhancement under adverse acoustic conditions.
Hsinyu Chang   +3 more
doaj   +3 more sources

Audio Visual Speaker Localization from EgoCentric Views [PDF]

open access: yesarXiv.org, 2023
The use of audio and visual modality for speaker localization has been well studied in the literature by exploiting their complementary characteristics. However, most previous works employ the setting of static sensors mounted at fixed positions.
Jinzheng Zhao   +3 more
semanticscholar   +1 more source

Home - About - Disclaimer - Privacy