Results 131 to 140 of about 47,965 (148)

MedChat: a fully offline multimodal AI system for privacy-preserving clinical anamnesis. [PDF]

open access: yesFront Artif Intell
Ruhland JB   +4 more
europepmc   +1 more source

Impact of cisplatin dose, renal function, and other factors on audiometrically-assessed ototoxicity in more than 1400 adult-onset cancer survivors from The Platinum Study: a multicentre cohort study. [PDF]

open access: yesEClinicalMedicine
Sanchez VA   +19 more
europepmc   +1 more source

End-to-End Speech Recognition in Russian

2018
End-to-end speech recognition systems incorporating deep neural networks (DNNs) have achieved good results. We propose applying CTC (Connectionist Temporal Classification) models and attention-based encoder-decoder in automatic recognition of the Russian continuous speech.
Nikita Markovnikov   +2 more
openaire   +1 more source

End-to-end Korean Digits Speech Recognition

2019 International Conference on Information and Communication Technology Convergence (ICTC), 2019
The traditional speech recognition model consisting of an acoustic model and a language model is mainly used. Recently, an end-to-end speech recognition model consisting of a single integrated neural network model is being studied. This model has the advantage that it does not require a lot of training and it is easy to understand the structure of the ...
Jong-Hyuk Roh   +3 more
openaire   +1 more source

End-to-End Speech Recognition in Agglutinative Languages

2020
This paper considers end-to-end speech recognition systems based on deep neural networks (DNN). The studies used different types of neural networks, CTC model and attention-based encoder-decoder models. As a result of the study, it was proved that the CTC model works without language models directly for agglutinative languages, but the best is ResNet ...
Orken Mamyrbayev   +4 more
openaire   +1 more source

End-to-End Speech Recognition

2019
In Chap. 8, we aimed to create an ASR system by dividing the fundamental equation $$\displaystyle W^* = \operatorname *{argmax}_{W \in V^*} P(W|X) $$ into an acoustic model, lexicon model, and language model by using Bayes’ theorem. This approach relies heavily on the use of the conditional independence assumption and separate optimization ...
Uday Kamath, John Liu, James Whitaker
openaire   +1 more source

Parameter Uncertainty for End-to-end Speech Recognition

ICASSP 2019 - 2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), 2019
Recent work on neural networks with probabilistic parameters has shown that parameter uncertainty improves network regularization. Parameter-specific signal-to-noise ratio (SNR) levels derived from parameter distributions were further found to have high correlations with task importance.
Stefan Braun 0005, Shih-Chii Liu
openaire   +2 more sources

Home - About - Disclaimer - Privacy