Results 241 to 250 of about 2,081,168 (298)
Some of the next articles are maybe not open access.

Leveraging Sound Localization to Improve Continuous Speaker Separation

IEEE International Conference on Acoustics, Speech, and Signal Processing
Continuous speaker separation aims to separate overlapping speakers in real-world environments like meetings, but it often falls short in isolating speech segments of a single speaker.
Deliang Wang   +2 more
exaly   +2 more sources

Performance of speaker localization using microphone array

International Journal of Speech Technology, 2016
Speaker localization is a technique to locate and track an active speaker from multiple acoustic sources using microphone array. Microphone array is used to improve the speech quality of recorded speech signal in meeting room and other places. In this work, the time delay estimation between source and each microphone is calculated using a localization ...
Palanivel S   +2 more
exaly   +2 more sources

Robust Time-Delay Estimation for Speaker Localization Using Mutual Information Among Multiple Microphone Signals

IEEE Sensors Journal, 2023
Time-delay estimation (TDE) algorithms for speaker localization usually suffer from adverse effects of background noise and reverberation. The multichannel cross correlation coefficient (MCCC) algorithm exploits spatial redundancy among multiple ...
Juping Wang   +4 more
semanticscholar   +1 more source

Time Delay Estimation for Speaker Localization Using CNN-Based Parametrized GCC-PHAT Features

Interspeech, 2021
We propose a time delay estimation (TDE) method for speaker localization based on parametrized generalized cross-correlation phase transform (PGCC-PHAT) functions and convolutional neural networks (CNNs).
D. Salvati, C. Drioli, G. Foresti
semanticscholar   +1 more source

Deep learning based stage-wise two-dimensional speaker localization with large ad-hoc microphone arrays

Speech Communication, 2022
While deep-learning-based speaker localization has shown advantages in challenging acoustic environments, it often yields only direction-of-arrival (DOA) cues rather than precise two-dimensional (2D) coordinates.
Shupei Liu   +6 more
semanticscholar   +1 more source

Robust Indoor Speaker Localization in the Circular Harmonic Domain

IEEE transactions on industrial electronics (1982. Print), 2021
In this article, modal signal processing using circular sensor arrays has been shown to be an attractive way for speaker localization, due to the fact that it enables high flexibility in sound field analysis, which inherently supports wideband acoustic ...
Kunkun SongGong, Huawei Chen
semanticscholar   +1 more source

CNN-Based Processing of Acoustic and Radio Frequency Signals for Speaker Localization from MAVs

Interspeech, 2021
A novel speaker localization algorithm from micro aerial vehicles (MAVs) is investigated. It introduces a joint direction of arrival (DOA) and distance prediction method based on processing and fusion of the multi-channel speech data with radio frequency
Andrea Toma   +3 more
semanticscholar   +1 more source

Multiple Speaker Localization using Mixture of Gaussian Model with Manifold-based Centroids

European Signal Processing Conference, 2021
A data-driven approach for multiple speakers localization in reverberant enclosures is presented. The approach combines semi-supervised learning on multiple manifolds with unsupervised maximum likelihood estimation. The relative transfer functions (RTFs)
Avital Bross   +2 more
semanticscholar   +1 more source

Locality preserving speaker clustering

2009 IEEE International Conference on Multimedia and Expo, 2009
In this paper, we propose an efficient speaker clustering approach based on a locality preserving linear projective mapping in the Gaussian mixture model (GMM) mean supervector space. While the GMM mean supervector has turned out to be an effective representation of speakers, its dimensionality is usually very high.
Stephen M. Chu   +2 more
openaire   +1 more source

Localization-Driven Speech Enhancement in Noisy Multi-Speaker Hospital Environments Using Deep Learning and Meta Learning

IEEE/ACM Transactions on Audio Speech and Language Processing, 2023
This work addresses the problem of 3D-localizing and enhancing the speech of one main speaker in noisy multi-speaker hospital environments using a multi-channel microphone array.
Mahdi Barhoush   +4 more
semanticscholar   +1 more source

Home - About - Disclaimer - Privacy