A Dual-Stage Perceptual-Harmonic Hybrid Estimator for Speech Enhancement
This paper proposes a hybrid speech enhancement estimator that integrates the Perceptually-motivated Karhunen–Loève Transform (PKLT) with the Dual-Masking Harmonic-based (DMH) algorithm in a unified framework termed PKDMH.
Sally Taha Yousif , Basheera M. Mahmmod
doaj +1 more source
Comparison of MSE, SSIM, PESQ, MOS, CPU time (s), and memory (Kb).
Comparison of MSE, SSIM, PESQ, MOS, CPU time (s), and memory (Kb).
Zaid Ameen Abduljabbar (17729630) +7 more
core +1 more source
Noise Reduction from Speech Signal based on Wavelet Transform and Kullback-Leibler Divergence
A new method for speech enhancement based on Kullback-Leibler (K-L) divergence has been presented in this paper. First, the algorithm performs wavelet-packet transform to noisy speech and decomposes it into sub-bands; then we apply a threshold on ...
Shima Tabibian, Ahmad Akbari
doaj
Few-Shot Transfer for Speech Enhancement Using SEGAN with Stability Guardrails
High-quality speech communication is often compromised by background noise, reducing intelligibility and perceived quality.We investigate data-efficient few-shot transfer of the speech enhancement generative adversarial network (SEGAN) to a new noise ...
Rubi Sharma, Firos A.
doaj +1 more source
The investigation of frame disturbance (FD) in perceptual evaluation speech quality (PESQ) as a perceptual metric [PDF]
© 2006-2015 Asian Research Publishing Network (ARPN). Satisfying customers' needs economically is one of the important aspects in mobile communication industry.
A. Jusoh (23550760) +4 more
core +1 more source
Speech Enhancement Based on Dual-Path Cross-Parallel Conformer Network
The speech quality can be reduced by long-term and short-term noise, and frequency-domain single-path speech enhancement methods suffer from phase missing.
Qing Zhao +3 more
doaj +1 more source
The average STOI and PESQ improvements (STOIi and PESQi) of deep learning models over noisy speech.
The average STOI and PESQ improvements (STOIi and PESQi) of deep learning models over noisy speech.
Amil Daraz (9667034) +3 more
core +1 more source
Prevalência de maloclusões e associação com hábitos de sucção em pré-escolares do município de Florianópolis [PDF]
TCC (graduação) - Universidade Federal de Santa Catarina. Centro de Ciências da Saúde. Odontologia.As maloclusões são anomalias de desenvolvimento que geram alterações nos arcos dentários.
Santos, Júlia Gonçalves dos
core
Single-Channel Speech Enhancement Algorithm Based on ME-MGCRN in Low Signal-to-Noise Scenario
In low signal-to-noise ratio (SNR) conditions, to address the problem of poor speech enhancement effect of traditional neural networks, this paper combines Convolution Recurrent Neural Network (CRN) with Gated Linear Units (GLU) to extract speech ...
Chaofeng Lan +6 more
doaj +1 more source
PESQ and STOI result for LibriSPeech dataset.
Long short-term memory (LSTM) has been effectively used to represent sequential data in recent years. However, LSTM still struggles with capturing the long-term temporal dependencies.
Amil Daraz (9667034) +3 more
core +1 more source

