Results 11 to 20 of about 981,169 (268)

wav2vec2-based Speech Rating System for Children with Speech Sound Disorder

open access: yes, 2022
The computational resources were provided by Aalto ScienceIT. This work was supported by NordForsk through the funding to Technology-enhanced foreign and second-language learning of Nordic languages, project number 103893.Speaking is a fundamental way of
Getman, Yaroslav   +8 more
core   +1 more source

GuidedMix: An on‐the‐fly data augmentation approach for robust speaker recognition system

open access: yesElectronics Letters, 2022
Data augmentation is an essential technique for building a high‐robustness speaker recognition system. this letter proposes a novel on‐the‐fly data augmentation strategy called GuidedMix.
Runqiu Xiao   +4 more
doaj   +1 more source

Restricted Boltzmann Machine-Based Approaches for Link Prediction in Dynamic Networks

open access: yesIEEE Access, 2018
Link prediction in dynamic networks aims to predict edges according to historical linkage status. It is inherently difficult because of the linear/non-linear transformation of underlying structures.
Taisong Li   +4 more
doaj   +1 more source

A Formant Modification Method for Improved ASR of Children’s Speech

open access: yes, 2022
Differences in acoustic characteristics between children’s and adults’ speech degrade performance of automatic speech recognition systems when systems trained using adults’ speech are used to recognize children’s speech.
Alku, Paavo   +3 more
core   +1 more source

Applying deep matching networks to Chinese medical question answering: a study and a dataset

open access: yesBMC Medical Informatics and Decision Making, 2019
Background Medical and clinical question answering (QA) is highly concerned by researchers recently. Though there are remarkable advances in this field, the development in Chinese medical domain is relatively backward.
Junqing He, Mingming Fu, Manshu Tu
doaj   +1 more source

Parallel global convolutional network for semantic image segmentation

open access: yesIET Image Processing, 2021
In this paper, a novel convolutional neural network for fast semantic segmentation is presented. Deep convolutional neural networks have achieved great progress in the task of vision scene understanding.
Xing Bai, Jun Zhou
doaj   +1 more source

Comparison of glottal closure instants detection algorithms for emotional speech

open access: yes, 2020
avaa julkaisu, kun artikkeli saatavillaIn production of voiced speech, epochs or glottal closure instants (GCIs) refer to the instants of significant excitation of the vocal tract.
Sudarsana Reddy Kadiri   +5 more
core   +1 more source

Improving Hybrid CTC/Attention Architecture with Time-Restricted Self-Attention CTC for End-to-End Speech Recognition

open access: yesApplied Sciences, 2019
As demonstrated in hybrid connectionist temporal classification (CTC)/Attention architecture, joint training with a CTC objective is very effective to solve the misalignment problem existing in the attention-based end-to-end automatic speech recognition (
Long Wu, Ta Li, Li Wang, Yonghong Yan
doaj   +1 more source

A Complementary Effect in Active Control of Powertrain and Road Noise in the Vehicle Interior

open access: yesIEEE Access, 2022
This study shows that a concurrent active noise control strategy for engine harmonics and road noise has a complementary effect. In particular, we found that engine booming noise is additionally attenuated when road noise control is concurrently used ...
Seonghyeon Kim, M. Ercan Altinsoy
doaj   +1 more source

Comprehensive Active Control of Booming Noise Inside a Vehicle Caused by the Engine and the Driveline

open access: yesIEEE Access, 2022
This study presents comprehensive active cancellation of booming noise caused by the engine and the driveline inside a passenger car. In modern noise control systems for vehicles, booming noise caused by engine harmonics could be effectively suppressed ...
Seonghyeon Kim, M. Ercan Altinsoy
doaj   +1 more source

Home - About - Disclaimer - Privacy