Results 41 to 50 of about 3,698 (171)
On Losses, Pauses, Jumps, and the Wideband E-Model
There is an increasing interest in upgrading the E-Model, a parametric tool for speech quality estimation, to the wideband and super-wideband contexts. The main motivation behind this has been to quantify the quality gain lent by various new codecs and ...
Muhammad Adil Raja +2 more
doaj +1 more source
Applications of Artificial Intelligence in Neurological Voice Disorders
ABSTRACT Neurological voice disorders, such as Parkinson's disease, laryngeal dystonia, and stroke‐induced dysarthria, significantly impact speech production and communication. Traditional diagnostic methods rely on subjective assessment, whereas artificial intelligence (AI) offers objective, noninvasive, and scalable solutions for voice analysis. This
Dongren Yao +2 more
wiley +1 more source
ABSTRACT Interventions to enhance positive parenting practices have become a cornerstone of many Western child welfare services. Parental mental health is a crucial factor that influences parenting practices and styles. However, research on the associations between mental health and parenting among parents involved with child welfare services is scarce.
Jannike Kaasbøll +2 more
wiley +1 more source
Lightweight Real-Time Recurrent Models for Speech Enhancement and Automatic Speech Recognition.
Traditional recurrent neural networks (RNNs) encounter difficulty in capturing long-term temporal dependencies. However, lightweight recurrent models for speech enhancement are important to improve noisy speech, while being computationally efficient and
Sami Dhahbi +6 more
doaj +1 more source
Assessing Omnichannel Service Quality in the Public Sector
ABSTRACT Services in the public sector are often provided as a hybrid combination of both digital (e.g., online registration) and physical (e.g., offline appointment) service channels, which can be referred to as omnichannel services. There is a lack of instruments in the literature that can measure the perceived quality of public sector omnichannel ...
Fabian Walke, Till J. Winkler
wiley +1 more source
Self-supervised learning (SSL) models have achieved remarkable success in speaker verification tasks, yet their robustness to real-world audio degradation remains insufficiently characterized.
Ajan Ahmed, Masudul H. Imtiaz
doaj +1 more source
A neural coding method based on feature sensing
It has been shown that human neural coding is characterized by non‐linearity and that different neurons are more responsive to stimuli with specific features. Based on this, we propose a neural coding method based on feature sensing (feature sensing neural coding; FSNC). Taking speech signals as an example to simulate the process of auditory formation,
Dongbin He, Aiqun Hu, Kaiwen Sheng
wiley +1 more source
The PESQ improvements (PESQi) in background noises.
The PESQ improvements (PESQi) in background noises.
Amil Daraz (9667034) +3 more
core +1 more source
In this paper, a multi-task learning U-shaped neural network (MTU-Net) is proposed and applied to single-channel speech enhancement (SE). The proposed MTU-based SE method estimates an ideal binary mask (IBM) or an ideal ratio mask (IRM) by extending the ...
Geon Woo Lee, Hong Kook Kim
doaj +1 more source
This article proposes a distributed minimum variance distortionless response (MVDR) beamformer for speech enhancement, where the network‐wide steering vector used for generating MVDR is estimated in a distributed way. This method is iteratively performed, and in each iteration, the steering matrix and the beamformer coefficient are simultaneously ...
Xingyue Cui, Rui Wang
wiley +1 more source

