Results 31 to 40 of about 2,081,168 (298)

A Coherent Wideband Acoustic Source Localization Using a Uniform Circular Array

open access: yesSensors, 2023
In modern applications such as robotics, autonomous vehicles, and speaker localization, the computational power for sound source localization applications can be limited when other functionalities get more complex.
Meng Jiang   +4 more
doaj   +1 more source

Leveraging Visual Supervision for Array-Based Active Speaker Detection and Localization [PDF]

open access: yesIEEE/ACM Transactions on Audio Speech and Language Processing, 2023
Conventional audio-visual approaches for active speaker detection (ASD) typically rely on visually pre-extracted face tracks and the corresponding single-channel audio to find the speaker in a video.
Davide Berghi, Philip J. B. Jackson
semanticscholar   +1 more source

Dynamically localizing multiple speakers based on the time-frequency domain

open access: yesEURASIP Journal on Audio, Speech, and Music Processing, 2021
In this study, we present a deep neural network-based online multi-speaker localization algorithm based on a multi-microphone array. Following the W-disjoint orthogonality principle in the spectral domain, time-frequency (TF) bin is dominated by a single
Hodaya Hammer   +3 more
doaj   +1 more source

Effects of Binaural Spatialization in Wireless Microphone Systems for Hearing Aids on Normal-Hearing and Hearing-Impaired Listeners

open access: yesTrends in Hearing, 2018
Little is known about the perception of artificial spatial hearing by hearing-impaired subjects. The purpose of this study was to investigate how listeners with hearing disorders perceived the effect of a spatialization feature designed for wireless ...
Gilles Courtois   +4 more
doaj   +1 more source

Audio Inputs for Active Speaker Detection and Localization Via Microphone Array [PDF]

open access: yesIEEE Workshop on Applications of Signal Processing to Audio and Acoustics, 2023
This study considers the problem of detecting and locating an active talker’s horizontal position from multichannel audio captured by a microphone array. We refer to this as active speaker detection and localization (ASDL).
Davide Berghi, P. Jackson
semanticscholar   +1 more source

Localization of Directional Sound Sources Supported by A Priori Information of the Acoustic Environment

open access: yesEURASIP Journal on Advances in Signal Processing, 2007
Speaker localization with microphone arrays has received significant attention in the past decade as a means for automated speaker tracking of individuals in a closed space for videoconferencing systems, directed speech capture systems, and surveillance ...
András Radványi   +1 more
doaj   +1 more source

Cross Modal Video Representations for Weakly Supervised Active Speaker Localization [PDF]

open access: yesIEEE transactions on multimedia, 2020
An objective understanding of media depictions, such as inclusive portrayals of how much someone is heard and seen on screen such as in film and television, requires the machines to discern automatically who, when, how, and where someone is talking, and ...
Rahul Sharma   +2 more
semanticscholar   +1 more source

Migration-induced Variation in the Communicative Space

open access: yesGragoatá, 2017
Migration is one of the fundamental and widespread causes of linguistic contact and variation induced by contact. Although contact-induced linguistic variation is not uncommon, it requires explanation within a model that allows the linguistic data to be ...
Thomas Krefeld
doaj   +1 more source

Fast Sound Source Localization Based on SRP-PHAT Using Density Peaks Clustering

open access: yesApplied Sciences, 2021
Sound source localization has been increasingly used recently. Among the existing techniques of sound source localization, the steered response power–phase transform (SRP-PHAT) exhibits considerable advantages regarding anti-noise and anti-reverberation.
De-Bing Zhuo, Hui Cao
doaj   +1 more source

Audio Segmentation and Speaker Localization in Meeting Videos [PDF]

open access: yes18th International Conference on Pattern Recognition (ICPR'06), 2006
Segmenting different individuals in a group meeting and their speech is an important first step for various tasks such as meeting transcription, automatic camera panning, multimedia retrieval and monologue detection. In this effort, given a meeting room video, we attempt to segment individual person’s speech and localize them in the video, based on ...
Himanshu Vajaria   +4 more
openaire   +1 more source

Home - About - Disclaimer - Privacy