Multilingual Meta-Transfer Learning for Low-Resource Speech Recognition
This paper proposes a novel meta-transfer learning method to improve automatic speech recognition (ASR) performance in low-resource languages. Nowadays, we are witnessing high interest in low-resource ASR tasks aiming at delivering feasible and reliable ...
Rui Zhou +4 more
doaj +1 more source
The transformer is a Deep Learning (DL) model that revolutionized language processing with its self-attention mechanism, enabling parallel processing and improving model efficiency, which dramatically reshaped the landscape of speech recognition ...
Yousef O. Sharrab +4 more
doaj +1 more source
OS-Denseformer: A Lightweight End-to-End Noise-Robust Method for Chinese Speech Recognition
Automatic speech recognition (ASR) technology faces the dual challenges of model complexity and noise robustness when deployed on terminal devices (e.g., mobile devices, embedded systems). To meet the demand for lightweight and high-performance models in
Shiqi Que +3 more
doaj +1 more source
An Input-synchronous Blockwise Decoding Algorithm for CTC-AED Speech Recognition
Automatic speech recognition (ASR) systems for real-life scenarios are required to process audio streams of arbitrary length with stable accuracy under limited computational resources.
Iurii Lezhenin, Natalia Bogach
doaj +1 more source
An End-To-End Speech Recognition Model for the North Shaanxi Dialect: Design and Evaluation. [PDF]
Qin Y, Yu F.
europepmc +1 more source
Automatic Speech Recognition and Acoustic Analysis for Dysarthria Assessment in Telerehabilitation: User-Centered Design and Usability Study. [PDF]
Vinet P +9 more
europepmc +1 more source
Adaptive Phoneme State Learning Architecture for Enhanced Speech Recognition Using Backpropagation Neural Network and Hidden Markov Model. [PDF]
Siddalingappa R +8 more
europepmc +1 more source
AI-based services for inclusive language learning in immersive XR environments: Speech translation, and sign language integration. [PDF]
Tantaroudas ND +3 more
europepmc +1 more source
Recovering Speech from Vibrations: Principles and Algorithms in Radar and Laser Sensing. [PDF]
Bederov E, Berdugo B, Cohen I.
europepmc +1 more source
Bidirectional Kazakh Sign Language prosody-aware translation using computer vision and speech recognition techniques. [PDF]
Zhassuzak M +4 more
europepmc +1 more source

