Exploring Prosodic Features Modelling for Secondary Emotions Needed for Empathetic Speech Synthesis [PDF]
A low-resource emotional speech synthesis system for empathetic speech synthesis based on modelling prosody features is presented here. Secondary emotions, identified to be needed for empathetic speech, are modelled and synthesised in this investigation.
Jesin James +3 more
doaj +2 more sources
Time-domain concatenative text-to-speech synthesis. [PDF]
A concatenation framework for time-domain concatenative speech synthesis (TDCSS) is presented and evaluated. In this framework, speech segments are extracted from CV, VC, CVC and CC waveforms, and abutted.
Vine, Daniel Samuel Gordon
core +9 more sources
Prosodic Prominence and Boundaries in Sequence-to-Sequence Speech Synthesis [PDF]
Recent advances in deep learning methods have elevated synthetic speech quality to human level, and the field is now moving towards addressing prosodic variation in synthetic speech.Despite successes in this effort, the state-of-the-art systems fall ...
Juraj Šimko +7 more
core +1 more source
Analysis of speech prosody using WaveNet embeddings : The Lombard effect [PDF]
We present a novel methodology for speech prosody research based on the analysis of embeddings used to condition a convolutional WaveNet speech synthesis system.
Juraj Šimko +5 more
core +1 more source
Chinese personalised text‐to‐speech synthesis for robot human–machine interaction
Speech interaction is an important means of robot interaction. With the rapid development of deep learning, end‐to‐end speech synthesis methods based on this technique have gradually become mainstream.
Bao Pang +5 more
doaj +1 more source
An Emotion Speech Synthesis Method Based on VITS
People and things can be connected through the Internet of Things (IoT), and speech synthesis is one of the key technologies. At this stage, end-to-end speech synthesis systems are capable of synthesizing relatively realistic human voices, but the ...
Wei Zhao, Zheng Yang
doaj +1 more source
Computer-Implemented Articulatory Models for Speech Production: A Review
Modeling speech production and speech articulation is still an evolving research topic. Some current core questions are: What is the underlying (neural) organization for controlling speech articulation?
Bernd J. Kröger
doaj +1 more source
Unsupervised cross-lingual speaker adaptation for HMM-based speech synthesis using two-pass decision tree construction [PDF]
This paper demonstrates how unsupervised cross-lingual adaptation of HMM-based speech synthesis models may be performed without explicit knowledge of the adaptation data language.
Teemu Hirsimaki +5 more
core +2 more sources
Evaluating Prosodic Processing for Incremental Speech Synthesis [PDF]
Baumann T, Schlangen D. Evaluating Prosodic Processing for Incremental Speech Synthesis.
Schlangen, David +4 more
core +1 more source
A survey of expressive speech synthesis
Speech synthesis is a hot research topic in the field of speech, language and machine learning, which aims to synthesize understandable and natural speech for a given text.It has a wide range of applications in industry.One of the goals of speech ...
Haobin TANG +4 more
doaj

