Results 11 to 20 of about 83,033 (339)

Fully-Hierarchical Fine-Grained Prosody Modeling For Interpretable Speech Synthesis [PDF]

open access: yesIEEE International Conference on Acoustics, Speech, and Signal Processing, 2020
This paper proposes a hierarchical, fine-grained and interpretable latent variable model for prosody based on the Tacotron 2 text-to-speech model. It achieves multi-resolution modeling of prosody by conditioning finer level representations on coarser ...
Guangzhi Sun   +5 more
semanticscholar   +1 more source

Collocation, semantic prosody and near synonymy: A cross-linguistic perspective. [PDF]

open access: yes, 2006
This paper explores the collocational behaviour and semantic prosody of near synonyms from a cross-linguistic perspective. The importance of these concepts to language learning is well recognized.
Xiao, R. Z., McEnery, A. M.
core   +5 more sources

Generating Diverse and Natural Text-to-Speech Samples Using a Quantized Fine-Grained VAE and Autoregressive Prosody Prior [PDF]

open access: yesIEEE International Conference on Acoustics, Speech, and Signal Processing, 2020
Recent neural text-to-speech (TTS) models with fine-grained latent features enable precise control of the prosody of synthesized speech. Such models typically incorporate a fine-grained variational autoencoder (VAE) structure, extracting latent features ...
Guangzhi Sun   +7 more
semanticscholar   +1 more source

Prosody and Corpora

open access: yesCadernos de Linguística, 2021
This paper focuses on the experience of spoken corpora compilation and discusses the relevance of prosody in this type of endeavor, as well as in the study of spoken language in its several possibilities. Through the voices of scholars associated with four different projects (CorpAfroAs, Mohawk Corpus, LABLITA, C-ORAL-BRASIL), the steps considered of ...
Mello, Heliana   +4 more
openaire   +4 more sources

Speech Prosody in Mental Disorders

open access: yesAnnual Review of Linguistics, 2022
In response to uncovering brain mechanisms underlying vocal communication and searching for biomarkers for mental illnesses, speech prosody has been increasingly studied in recent years in multiple disciplines, including psycholinguistics.
Hongwei Ding, Yang Zhang
semanticscholar   +1 more source

CopyCat: Many-to-Many Fine-Grained Prosody Transfer for Neural Text-to-Speech [PDF]

open access: yesInterspeech, 2020
Prosody Transfer (PT) is a technique that aims to use the prosody from a source audio as a reference while synthesising speech. Fine-grained PT aims at capturing prosodic aspects like rhythm, emphasis, melody, duration, and loudness, from a source audio ...
S. Karlapati   +5 more
semanticscholar   +1 more source

Mark my words: tone of voice changes affective word representations in memory. [PDF]

open access: yesPLoS ONE, 2010
The present study explored the effect of speaker prosody on the representation of words in memory. To this end, participants were presented with a series of words and asked to remember the words for a subsequent recognition test. During study, words were
Annett Schirmer
doaj   +1 more source

Prosodie et prosodie musicale

open access: yesLa Bretagne linguistique, 1993
Comment la langue et son accentuation s’adaptent-elles aux contraintes de la mélodie ? De ces deux éléments, qu’est-ce qui prime dans l’élaboration d’un chant ? La mise en évidence de la relation entre l’accentuation des paroles et celle de la mélodie est le but de la prosodie musicale.
Laurent, Donatien, Goyat, Gilles
openaire   +5 more sources

“Textual Prosody” Can Change Impressions of Reading in People With Normal Hearing and Hearing Loss

open access: yesFrontiers in Psychology, 2020
Recently, dynamic text presentation, such as scrolling text, has been widely used. Texts are often presented at constant timing and speed in conventional dynamic text presentation.
Miki Uetsuki   +2 more
doaj   +1 more source

Hierarchical Prosody Modeling for Non-Autoregressive Speech Synthesis [PDF]

open access: yesSpoken Language Technology Workshop, 2020
Prosody modeling is an essential component in modern text-to-speech (TTS) frameworks. By explicitly providing prosody features to the TTS model, the style of synthesized utterances can thus be controlled.
C. Chien, Hung-yi Lee
semanticscholar   +1 more source

Home - About - Disclaimer - Privacy