Results 11 to 20 of about 178,826,221 (80)
wav2vec2-based Speech Rating System for Children with Speech Sound Disorder
The computational resources were provided by Aalto ScienceIT. This work was supported by NordForsk through the funding to Technology-enhanced foreign and second-language learning of Nordic languages, project number 103893.Speaking is a fundamental way of
Getman, Yaroslav +8 more
core +1 more source
Prosodic Prominence and Boundaries in Sequence-to-Sequence Speech Synthesis [PDF]
Recent advances in deep learning methods have elevated synthetic speech quality to human level, and the field is now moving towards addressing prosodic variation in synthetic speech.Despite successes in this effort, the state-of-the-art systems fall ...
Juraj Šimko +7 more
core +1 more source
Analysis of speech prosody using WaveNet embeddings : The Lombard effect [PDF]
We present a novel methodology for speech prosody research based on the analysis of embeddings used to condition a convolutional WaveNet speech synthesis system.
Juraj Šimko +5 more
core +1 more source
Purpose: This study examined speech breathing during two connected speech tasks in children with cerebral palsy (CP) and typically developing (TD) peers.
Darling-White, Meghan, Kovacs, Sydney
core +1 more source
Dysarthric speech classification using glottal features computed from non-words, words and sentences
Dysarthria is a neuro-motor disorder resulting from the disruption of normal activity in speech production leading to slow, slurred and imprecise (low intelligible) speech.
Narendra N P +3 more
core +1 more source
Sound Privacy: A Conversational Speech Corpus for Quantifying the Experience of Privacy
With the growing popularity of social networks, cloud services and online applications, people are becoming concerned about the way companies store their data and the ways in which the data can be applied.
Das, Sneha +9 more
core +1 more source
Lombard speech synthesis using transfer learning in a Tacotron text-to-speech system
Currently, there is increasing interest to use sequence-to-sequence models in text-to-speech (TTS) synthesis with attention like that in Tacotron models.
Bajibabu Bollepalli +5 more
core +1 more source
Automatic classification of vocal intensity category from speech
Regulation of vocal intensity is a fundamental phenomenon in speech communication. Vocal intensity can be quantified using sound pressure level (SPL), which can be measured easily by recording a standard calibration signal with speech and by comparing ...
Laaksonen, Laura +3 more
core +1 more source
Speech coding is the most commonly used application of speech processing. Accumulated layers of improvements have however made codecs so complex that optimization of individual modules becomes increasingly difficult. This work introduces machine learning
Bäckström, Tom, Tom Bäckström
core +1 more source
The relationship between acoustic indices of speech motor control variability and other measures of speech performance in dysarthria [PDF]
Previous studies suggested that variability indices based on information extracted from the acoustic signal are potentially useful in assessing dysarthric speech.
Van Brenk, Frits, Lowit, Anja
core +3 more sources

