Results 11 to 20 of about 10,689,404 (277)

Speech Level Shift in the Domain of the Elderly Care in Japan by the Indonesian Caregivers [PDF]

open access: yesJournal of Language Teaching and Research, 2019
This study aimed at identifying the phenomenon of speech level shift in the domain of the elderly¬ care in Japan and finding out the rationale of the shift. Speech level in Japanese language is defined as the formality and politeness level in an interaction, which in lingustics is related to the use of formal or non-formal forms at the end of an ...
Putu Dewi Merlyna Yuda Pramesti   +3 more
openaire   +2 more sources

Prosodic Prominence and Boundaries in Sequence-to-Sequence Speech Synthesis [PDF]

open access: yes, 2020
Recent advances in deep learning methods have elevated synthetic speech quality to human level, and the field is now moving towards addressing prosodic variation in synthetic speech.Despite successes in this effort, the state-of-the-art systems fall ...
Juraj Šimko   +7 more
core   +1 more source

Analysis of speech prosody using WaveNet embeddings : The Lombard effect [PDF]

open access: yes, 2020
We present a novel methodology for speech prosody research based on the analysis of embeddings used to condition a convolutional WaveNet speech synthesis system.
Juraj Šimko   +5 more
core   +1 more source

Phoneme and sentence-level ensembles for speech recognition [PDF]

open access: yes, 2011
We address the question of whether and how boosting and bagging can be used for speech recognition. In order to do this, we compare two different boosting schemes, one at the phoneme level and one at the utterance level, with a phoneme-level bagging ...
Dimitrakakis, Christos   +5 more
core   +1 more source

Tutorial: Speech Assessment for Multilingual Children Who Do Not Speak the Same Language(s) as the Speech-Language Pathologist.

open access: yes, 2023
PurposeThe aim of this tutorial is to support speech-language pathologists (SLPs) undertaking assessments of multilingual children with suspected speech sound disorders, particularly children who speak languages that are not shared with their SLP ...
McLeod, Sharynne   +2 more
core   +1 more source

Multi-level Channel Fusion Method for Speech Emotion Recognition [PDF]

open access: yesJisuanji kexue yu tansuo
Speech emotion recognition is key to the emotional cognitive ability of machine and crucial for improving human-machine interaction quality. However, most of the existing studies focus on the analysis of shallow features, ignoring the advantages of multi-
ZHANG Limin, LI Yang, CAI Hao, YAN Hao
doaj   +1 more source

wav2vec2-based Speech Rating System for Children with Speech Sound Disorder

open access: yes, 2022
The computational resources were provided by Aalto ScienceIT. This work was supported by NordForsk through the funding to Technology-enhanced foreign and second-language learning of Nordic languages, project number 103893.Speaking is a fundamental way of
Getman, Yaroslav   +8 more
core   +1 more source

Generalisation Gap of Keyword Spotters in a Cross-Speaker Low-Resource Scenario

open access: yesSensors, 2021
Models for keyword spotting in continuous recordings can significantly improve the experience of navigating vast libraries of audio recordings. In this paper, we describe the development of such a keyword spotting system detecting regions of interest in ...
Łukasz Lepak   +3 more
doaj   +1 more source

An Ergonomic Study on the ‘Morningness’ and ‘Eveningness’ of Call Center Agents and Its Effect on Cognitive Performance

open access: yesInternational Journal of Technology, 2017
The increasing adaptation of shiftwork in the Philippines and its reported adverse effects had encouraged research studies among Filipino workers. This study aims to identify the circadian clock behavior of the shiftworker and its relationship together ...
Alma Maria Jennifer Gutierrez   +3 more
doaj   +1 more source

High-Level Approaches to Confidence Estimation in Speech Recognition [PDF]

open access: yes, 2002
We describe some high-level approaches to estimating confidence scores for the words output by a speech recognizer. By "high-level" we mean that the proposed measures do not rely on decoder specific "side information" and so should find more general ...
Cox, Stephen, Dasmahapatra, Srinandan
core   +2 more sources

Home - About - Disclaimer - Privacy