Results 71 to 80 of about 457 (166)
Feature Integration Strategies for Neural Speaker Diarization in Conversational Telephone Speech
This paper addresses the challenge of optimizing end-to-end neural diarization systems for conversational telephone speech, focusing on diverse acoustic features beyond traditional Mel-filterbanks.
Juan Ignacio Alvarez-Trejos +2 more
doaj +1 more source
Speaker diarization, the process of segmenting audio into speaker-specific regions, plays a critical role in various speech technologies by determining "who spoke when" in a conversation.
Mohd Zulhafiz Rahim +2 more
doaj
Speaker Diarization: A Review of Objectives and Methods
Recorded audio often contains speech from multiple people in conversation. It is useful to label such signals with speaker turns, noting when each speaker is talking and identifying each speaker.
Douglas O’Shaughnessy
doaj +1 more source
Unsupervised Segmentation Methods of TV Contents
We present a generic algorithm to address various temporal segmentation topics of audiovisual contents such as speaker diarization, shot, or program segmentation. Based on a GLR approach, involving the ΔBIC criterion, this algorithm requires the value of
Elie El-Khoury +2 more
doaj +1 more source
Toward a Smartwatch-Based System for Automated Social Interaction Monitoring in Autism
Autism Spectrum Disorder (ASD) is a complex neurodevelopmental condition that affects social interactions and communication. Traditional assessment methods are often manual, subjective - lack ecological validity and scalability.
Srinivas Goud Kalagotla +4 more
doaj +1 more source
The task of speaker diarization is to answer the question "who spoke when?" In this paper, we present different clustering approaches which consist of Evolutionary Computation Algorithms (ECAs) such as Genetic Algorithm (GA), Particle Swarm Optimization (
Karim Dabbabi, Salah Hajji, Adnen Cherif
doaj +1 more source
Bayes factor based speaker segmentation for speaker diarization
This paper proposes the use of the Bayes Factor as a distance metric for speaker segmentation within a speaker diarization system. The proposed approach uses a pair of constant sized, sliding windows to compute the value of the Bayes Factor between the adjacent windows over the entire audio.
Wang, David +2 more
openaire +1 more source
A lightweight approach to real-time speaker diarization: from audio toward audio-visual data streams
This manuscript deals with the task of real-time speaker diarization (SD) for stream-wise data processing. Therefore, in contrast to most of the existing papers, it considers not only the accuracy but also the computational demands of individual ...
Frantisek Kynych +4 more
doaj +1 more source
Meta-learning with Latent Space Clustering in Generative Adversarial Network for Speaker Diarization. [PDF]
Pal M +7 more
europepmc +1 more source
Privacy-Oriented Manipulation of Speaker Representations
Speaker embeddings are ubiquitous, with applications ranging from speaker recognition and diarization to speech synthesis and voice anonymization. The amount of information held by these embeddings lends them versatility but also raises privacy concerns.
Francisco Teixeira +3 more
doaj +1 more source

