Results 241 to 250 of about 2,081,168 (298)
Some of the next articles are maybe not open access.
Leveraging Sound Localization to Improve Continuous Speaker Separation
IEEE International Conference on Acoustics, Speech, and Signal ProcessingContinuous speaker separation aims to separate overlapping speakers in real-world environments like meetings, but it often falls short in isolating speech segments of a single speaker.
Deliang Wang +2 more
exaly +2 more sources
Performance of speaker localization using microphone array
International Journal of Speech Technology, 2016Speaker localization is a technique to locate and track an active speaker from multiple acoustic sources using microphone array. Microphone array is used to improve the speech quality of recorded speech signal in meeting room and other places. In this work, the time delay estimation between source and each microphone is calculated using a localization ...
Palanivel S +2 more
exaly +2 more sources
IEEE Sensors Journal, 2023
Time-delay estimation (TDE) algorithms for speaker localization usually suffer from adverse effects of background noise and reverberation. The multichannel cross correlation coefficient (MCCC) algorithm exploits spatial redundancy among multiple ...
Juping Wang +4 more
semanticscholar +1 more source
Time-delay estimation (TDE) algorithms for speaker localization usually suffer from adverse effects of background noise and reverberation. The multichannel cross correlation coefficient (MCCC) algorithm exploits spatial redundancy among multiple ...
Juping Wang +4 more
semanticscholar +1 more source
Time Delay Estimation for Speaker Localization Using CNN-Based Parametrized GCC-PHAT Features
Interspeech, 2021We propose a time delay estimation (TDE) method for speaker localization based on parametrized generalized cross-correlation phase transform (PGCC-PHAT) functions and convolutional neural networks (CNNs).
D. Salvati, C. Drioli, G. Foresti
semanticscholar +1 more source
Speech Communication, 2022
While deep-learning-based speaker localization has shown advantages in challenging acoustic environments, it often yields only direction-of-arrival (DOA) cues rather than precise two-dimensional (2D) coordinates.
Shupei Liu +6 more
semanticscholar +1 more source
While deep-learning-based speaker localization has shown advantages in challenging acoustic environments, it often yields only direction-of-arrival (DOA) cues rather than precise two-dimensional (2D) coordinates.
Shupei Liu +6 more
semanticscholar +1 more source
Robust Indoor Speaker Localization in the Circular Harmonic Domain
IEEE transactions on industrial electronics (1982. Print), 2021In this article, modal signal processing using circular sensor arrays has been shown to be an attractive way for speaker localization, due to the fact that it enables high flexibility in sound field analysis, which inherently supports wideband acoustic ...
Kunkun SongGong, Huawei Chen
semanticscholar +1 more source
CNN-Based Processing of Acoustic and Radio Frequency Signals for Speaker Localization from MAVs
Interspeech, 2021A novel speaker localization algorithm from micro aerial vehicles (MAVs) is investigated. It introduces a joint direction of arrival (DOA) and distance prediction method based on processing and fusion of the multi-channel speech data with radio frequency
Andrea Toma +3 more
semanticscholar +1 more source
Multiple Speaker Localization using Mixture of Gaussian Model with Manifold-based Centroids
European Signal Processing Conference, 2021A data-driven approach for multiple speakers localization in reverberant enclosures is presented. The approach combines semi-supervised learning on multiple manifolds with unsupervised maximum likelihood estimation. The relative transfer functions (RTFs)
Avital Bross +2 more
semanticscholar +1 more source
Locality preserving speaker clustering
2009 IEEE International Conference on Multimedia and Expo, 2009In this paper, we propose an efficient speaker clustering approach based on a locality preserving linear projective mapping in the Gaussian mixture model (GMM) mean supervector space. While the GMM mean supervector has turned out to be an effective representation of speakers, its dimensionality is usually very high.
Stephen M. Chu +2 more
openaire +1 more source
IEEE/ACM Transactions on Audio Speech and Language Processing, 2023
This work addresses the problem of 3D-localizing and enhancing the speech of one main speaker in noisy multi-speaker hospital environments using a multi-channel microphone array.
Mahdi Barhoush +4 more
semanticscholar +1 more source
This work addresses the problem of 3D-localizing and enhancing the speech of one main speaker in noisy multi-speaker hospital environments using a multi-channel microphone array.
Mahdi Barhoush +4 more
semanticscholar +1 more source

