Results 71 to 80 of about 167,964,167 (121)

Multi-Task Learning U-Net for Single-Channel Speech Enhancement and Mask-Based Voice Activity Detection

open access: yesApplied Sciences, 2020
In this paper, a multi-task learning U-shaped neural network (MTU-Net) is proposed and applied to single-channel speech enhancement (SE). The proposed MTU-based SE method estimates an ideal binary mask (IBM) or an ideal ratio mask (IRM) by extending the ...
Geon Woo Lee, Hong Kook Kim
doaj   +1 more source

Acoustic echo cancellation based on two‐stage BLSTM

open access: yesElectronics Letters, Volume 60, Issue 7, April 2024.
A two‐stage bidirectional long short term memory (TS‐BLSTM) framework, incorporating multi‐head self‐attention mechanisms after each BLSTM block, is introduced. The BLSTM blocks are utilized to aggregate magnitude spectrogram information, modelling both time and frequency dependencies.
Zhiwei Niu   +3 more
wiley   +1 more source

EVALUATION OF SPEECH ENHANCEMENT IN NOISY CONDITIONS USING A SPECTRAL SUBTRACTION AND LINEAR PREDICTION COMBINATION [PDF]

open access: yesICTACT Journal on Communication Technology, 2012
Improving quality and intelligibility of speech signals in mobile devices has been studied with great interest in the past. Speech information in communication channels is usually corrupted by additive acoustic noise, reverberation or channel noise. This
M. Thirumarai Chellapandi, P. Kabilan
doaj  

Speech Enhancement Using Joint DNN‐NMF Model Learned with Multi‐Objective Frequency Differential Spectrum Loss Function

open access: yesIET Signal Processing, Volume 2024, Issue 1, 2024.
We propose a multi‐objective joint model of non‐negative matrix factorization (NMF) and deep neural network (DNN) with a new loss function for speech enhancement. The proposed loss function (LMOFD) is a weighted combination of a frequency differential spectrum mean squared error (MSE)‐based loss function (LFD) and a multi‐objective MSE loss function ...
Matin Pashaian   +2 more
wiley   +1 more source

Aspects of voice irregularity measurement in connected speech [PDF]

open access: yes, 2009
Applications of the use of connected speech material for the objective assessment of two primary physical aspects of voice quality are described and discussed.
Fourcin, A.
core  

Speech Enhancement for Electrolarynx Devices Using M-RLS: Intelligibility Improvement and Low-Power Hardware Feasibility

open access: yesIEEE Access
An electrolarynx, a widely used assistive device for voice restoration for individuals with laryngeal impairments, often suffers from mechanical buzz and poor intelligibility, limiting its everyday applicability.
M. Madhushankara   +5 more
doaj   +1 more source

End-to-End Optimized Multi-Stage Vector Quantization of Spectral Envelopes for Speech and Audio Coding

open access: yes, 2021
Spectral envelope modeling is an instrumental part of speech and audio codecs, which can be used to enable efficient entropy coding of spectral components.
Bäckström, Tom   +3 more
core   +1 more source

Noise estimation based on optimal smoothing and minimum controlled through recursive averaging for speech enhancement

open access: yesIntelligent Systems with Applications
One of the most significant challenges in the real time speech processing applications is the elimination of noise in the corrupted speech data. This noise can significantly impact the efficacy/performance of the applications of speech processing ...
Raghudathesh G P   +3 more
doaj   +1 more source

Source Separation with One Ear: Proposition for an Anthropomorphic Approach

open access: yesEURASIP Journal on Advances in Signal Processing, 2005
We present an example of an anthropomorphic approach, in which auditory-based cues are combined with temporal correlation to implement a source separation system.
Ramin Pichevar, Jean Rouat
doaj   +1 more source

Home - About - Disclaimer - Privacy