Results 31 to 40 of about 457 (166)
TSUP Speaker Diarization System for Conversational Short-phrase Speaker Diarization Challenge
This paper describes the TSUP team's submission to the ISCSLP 2022 conversational short-phrase speaker diarization (CSSD) challenge which particularly focuses on short-phrase conversations with a new evaluation metric called conversational diarization error rate (CDER).
Bowen Pang 0002 +7 more
openaire +2 more sources
Comparison of Diarization Tools for Building Speaker Database
This paper compares open source diarization toolkits (LIUM, DiarTK, ALIZE-Lia_Ral), which were designed for extraction of speaker identity from audio records without any prior information about the analysed data.
Eva Kiktova, Jozef Juhar
doaj +1 more source
A Survey of Speaker Recognition: Fundamental Theories, Recognition Methods and Opportunities
Humans can identify a speaker by listening to their voice, over the telephone, or on any digital devices. Acquiring this congenital human competency, authentication technologies based on voice biometrics, such as automatic speaker recognition (ASR), have
Muhammad Mohsin Kabir +4 more
doaj +1 more source
The use of long-term features for GMM- and i-vector-based speaker diarization systems
Several factors contribute to the performance of speaker diarization systems. For instance, the appropriate selection of speech features is one of the key aspects that affect speaker diarization systems.
Abraham Woubie Zewoudie +2 more
doaj +1 more source
Speaker Diarization: A Review of Recent Research [PDF]
Speaker diarization is the task of determining "who spoke when?" in an audio or video recording that contains an unknown amount of speech and also an unknown number of speakers. Initially, it was proposed as a research topic related to automatic speech recognition, where speaker diarization serves as an upstream processing step.
Anguera, Xavier +5 more
openaire +2 more sources
Speaker diarization and linking of large corpora [PDF]
Performing speaker diarization of a collection of recordings, where speakers are uniquely identified across the database, is a challenging task. In this context, inter-session variability compensation and reasonable computation times are essential to be addressed.
Ferras, Marc, Bourlard, Hervé
openaire +1 more source
Global Research and Imaging Platform (GRIP): “FreeSurfer” for Digital Voice Processing [PDF]
Abstract Background Digital voice is increasingly being recognized as a less biased and more scalable approach for identifying those with cognitive impairment in the early stages of Alzheimer's disease (AD) and other related dementias (ADRD). Just as brain MRI scans were widely adopted as a tool for in vivo detection of AD/ADRD, FreeSurfer was ...
Karjadi C +13 more
europepmc +2 more sources
Analysis of transition cost and model parameters in speaker diarization for meetings
There has been little work in the literature on the speaker diarization of meetings with multiple distance microphones since the publications in 2012 related to the last National Institute of Standards (NIST) Rich Transcription Evaluation Campaign in ...
Beatriz Martínez-González +4 more
doaj +1 more source
Evidence-based facilitator strategies for enhancing social engagement in groups of older adults with ADRD. [PDF]
Abstract INTRODUCTION Alzheimer's disease and related dementias (ADRD) is exacerbated by social isolation. One potential intervention for improving the health outcomes of individuals with ADRD is by providing opportunities for socialization that are highly engaging.
McKinley J +4 more
europepmc +2 more sources
In this paper, we present a novel speaker diarization system for streaming on-device applications. In this system, we use a transformer transducer to detect the speaker turns, represent each speaker turn by a speaker embedding, then cluster these embeddings with constraints from the detected speaker turns.
Wei Xia +6 more
openaire +2 more sources

