Results 11 to 20 of about 10,689,404 (277)
Speech Level Shift in the Domain of the Elderly Care in Japan by the Indonesian Caregivers [PDF]
This study aimed at identifying the phenomenon of speech level shift in the domain of the elderly¬ care in Japan and finding out the rationale of the shift. Speech level in Japanese language is defined as the formality and politeness level in an interaction, which in lingustics is related to the use of formal or non-formal forms at the end of an ...
Putu Dewi Merlyna Yuda Pramesti +3 more
openaire +2 more sources
Prosodic Prominence and Boundaries in Sequence-to-Sequence Speech Synthesis [PDF]
Recent advances in deep learning methods have elevated synthetic speech quality to human level, and the field is now moving towards addressing prosodic variation in synthetic speech.Despite successes in this effort, the state-of-the-art systems fall ...
Juraj Šimko +7 more
core +1 more source
Analysis of speech prosody using WaveNet embeddings : The Lombard effect [PDF]
We present a novel methodology for speech prosody research based on the analysis of embeddings used to condition a convolutional WaveNet speech synthesis system.
Juraj Šimko +5 more
core +1 more source
Phoneme and sentence-level ensembles for speech recognition [PDF]
We address the question of whether and how boosting and bagging can be used for speech recognition. In order to do this, we compare two different boosting schemes, one at the phoneme level and one at the utterance level, with a phoneme-level bagging ...
Dimitrakakis, Christos +5 more
core +1 more source
PurposeThe aim of this tutorial is to support speech-language pathologists (SLPs) undertaking assessments of multilingual children with suspected speech sound disorders, particularly children who speak languages that are not shared with their SLP ...
McLeod, Sharynne +2 more
core +1 more source
Multi-level Channel Fusion Method for Speech Emotion Recognition [PDF]
Speech emotion recognition is key to the emotional cognitive ability of machine and crucial for improving human-machine interaction quality. However, most of the existing studies focus on the analysis of shallow features, ignoring the advantages of multi-
ZHANG Limin, LI Yang, CAI Hao, YAN Hao
doaj +1 more source
wav2vec2-based Speech Rating System for Children with Speech Sound Disorder
The computational resources were provided by Aalto ScienceIT. This work was supported by NordForsk through the funding to Technology-enhanced foreign and second-language learning of Nordic languages, project number 103893.Speaking is a fundamental way of
Getman, Yaroslav +8 more
core +1 more source
Generalisation Gap of Keyword Spotters in a Cross-Speaker Low-Resource Scenario
Models for keyword spotting in continuous recordings can significantly improve the experience of navigating vast libraries of audio recordings. In this paper, we describe the development of such a keyword spotting system detecting regions of interest in ...
Łukasz Lepak +3 more
doaj +1 more source
The increasing adaptation of shiftwork in the Philippines and its reported adverse effects had encouraged research studies among Filipino workers. This study aims to identify the circadian clock behavior of the shiftworker and its relationship together ...
Alma Maria Jennifer Gutierrez +3 more
doaj +1 more source
High-Level Approaches to Confidence Estimation in Speech Recognition [PDF]
We describe some high-level approaches to estimating confidence scores for the words output by a speech recognizer. By "high-level" we mean that the proposed measures do not rely on decoder specific "side information" and so should find more general ...
Cox, Stephen, Dasmahapatra, Srinandan
core +2 more sources

