Results 11 to 20 of about 981,169 (268)
wav2vec2-based Speech Rating System for Children with Speech Sound Disorder
The computational resources were provided by Aalto ScienceIT. This work was supported by NordForsk through the funding to Technology-enhanced foreign and second-language learning of Nordic languages, project number 103893.Speaking is a fundamental way of
Getman, Yaroslav +8 more
core +1 more source
GuidedMix: An on‐the‐fly data augmentation approach for robust speaker recognition system
Data augmentation is an essential technique for building a high‐robustness speaker recognition system. this letter proposes a novel on‐the‐fly data augmentation strategy called GuidedMix.
Runqiu Xiao +4 more
doaj +1 more source
Restricted Boltzmann Machine-Based Approaches for Link Prediction in Dynamic Networks
Link prediction in dynamic networks aims to predict edges according to historical linkage status. It is inherently difficult because of the linear/non-linear transformation of underlying structures.
Taisong Li +4 more
doaj +1 more source
A Formant Modification Method for Improved ASR of Children’s Speech
Differences in acoustic characteristics between children’s and adults’ speech degrade performance of automatic speech recognition systems when systems trained using adults’ speech are used to recognize children’s speech.
Alku, Paavo +3 more
core +1 more source
Applying deep matching networks to Chinese medical question answering: a study and a dataset
Background Medical and clinical question answering (QA) is highly concerned by researchers recently. Though there are remarkable advances in this field, the development in Chinese medical domain is relatively backward.
Junqing He, Mingming Fu, Manshu Tu
doaj +1 more source
Parallel global convolutional network for semantic image segmentation
In this paper, a novel convolutional neural network for fast semantic segmentation is presented. Deep convolutional neural networks have achieved great progress in the task of vision scene understanding.
Xing Bai, Jun Zhou
doaj +1 more source
Comparison of glottal closure instants detection algorithms for emotional speech
avaa julkaisu, kun artikkeli saatavillaIn production of voiced speech, epochs or glottal closure instants (GCIs) refer to the instants of significant excitation of the vocal tract.
Sudarsana Reddy Kadiri +5 more
core +1 more source
As demonstrated in hybrid connectionist temporal classification (CTC)/Attention architecture, joint training with a CTC objective is very effective to solve the misalignment problem existing in the attention-based end-to-end automatic speech recognition (
Long Wu, Ta Li, Li Wang, Yonghong Yan
doaj +1 more source
A Complementary Effect in Active Control of Powertrain and Road Noise in the Vehicle Interior
This study shows that a concurrent active noise control strategy for engine harmonics and road noise has a complementary effect. In particular, we found that engine booming noise is additionally attenuated when road noise control is concurrently used ...
Seonghyeon Kim, M. Ercan Altinsoy
doaj +1 more source
This study presents comprehensive active cancellation of booming noise caused by the engine and the driveline inside a passenger car. In modern noise control systems for vehicles, booming noise caused by engine harmonics could be effectively suppressed ...
Seonghyeon Kim, M. Ercan Altinsoy
doaj +1 more source

