Results 31 to 40 of about 20,181 (116)

Development of a System for Automatic Recognition of Speech

open access: yesAdvances in Electrical and Electronic Engineering, 2003
The article gives a review of a research on processing and automatic recognition of speech signals (ARR) at the Department of Telecommunications of the Faculty of Electrical Engineering, University of Zilina.
Roman Jarina, Michal Kuba
doaj  

Masked-speech Recognition Using Human and Synthetic Cloned Speech

open access: yesTrends in Hearing
Voice cloning is used to generate synthetic speech that mimics vocal characteristics of human talkers. This experiment used voice cloning to compare human and synthetic speech for intelligibility, human-likeness, and perceptual similarity, all tested in ...
Lauren Calandruccio   +3 more
doaj   +1 more source

Audio-Visual Speech Recognition Using MPEG-4 Compliant Visual Features

open access: yesEURASIP Journal on Advances in Signal Processing, 2002
We describe an audio-visual automatic continuous speech recognition system, which significantly improves speech recognition performance over a wide range of acoustic noise levels, as well as under clean audio conditions.
Aleksic Petar S   +3 more
doaj   +1 more source

Streaming End-to-End Target-Speaker Automatic Speech Recognition and Activity Detection

open access: yesIEEE Access, 2023
Automatic speech recognition of a target speaker in the presence of interfering speakers remains a challenging issue. One approach to tackle this problem is target-speaker speech recognition, which conditions the recognition process on an embedding that ...
Takafumi Moriya   +4 more
doaj   +1 more source

I Hear You Eat and Speak: Automatic Recognition of Eating Condition and Food Type, Use-Cases, and Impact on ASR Performance. [PDF]

open access: yesPLoS ONE, 2016
We propose a new recognition task in the area of computational paralinguistics: automatic recognition of eating conditions in speech, i. e., whether people are eating while speaking, and what they are eating.
Simone Hantke   +6 more
doaj   +1 more source

AudVowelConsNet: A phoneme-level based deep CNN architecture for clinical depression diagnosis

open access: yesMachine Learning with Applications, 2020
Depression is a common and serious mood disorder that negatively affects the patient’s capacity of functioning normally in daily tasks. Speech is proven to be a vigorous tool in depression diagnosis. Research in psychiatry concentrated on performing fine-
Muhammad Muzammel   +4 more
doaj   +1 more source

An Impact of Narrowband Speech Codec Mismatch on a Performance of GMM-UBM Speaker Recognition over Telecommunication Channel

open access: yesCommunications, 2016
The automatic identification of person's identity from their voice is a part of modern telecommunication services. In order to execute the identification task, speech signal has to be transmitted to a remote server.
Jozef Polacky, Peter Pocta, Roman Jarina
doaj   +1 more source

Speech recognition in adverse conditions by humans and machines [PDF]

open access: yesJASA Express Letters
In the development of automatic speech recognition systems, achieving human-like performance has been a long-held goal. Recent releases of large spoken language models have claimed to achieve such performance, although direct comparison to humans has ...
Chloe Patman, Eleanor Chodroff
doaj   +1 more source

Post-error Correction in Automatic Speech Recognition Using Discourse Information

open access: yesAdvances in Electrical and Computer Engineering, 2014
Overcoming speech recognition errors in the field of human computer interaction is important in ensuring a consistent user experience. This paper proposes a semantic-oriented post-processing approach for the correction of errors in speech recognition ...
KANG, S., KIM, J.-H., SEO, J.
doaj   +1 more source

Home - About - Disclaimer - Privacy