Results 101 to 110 of about 3,633,480 (246)

A Theoretically Consistent Method for Minimum Mean-Square Error Estimation of Mel-Frequency Cepstral Features

open access: yes, 2014
We propose a method for minimum mean-square error (MMSE) estimation of mel-frequency cepstral features for noise robust automatic speech recognition (ASR).
Tan, Zheng-Hua; id_orcid   +3 more
core   +1 more source

Mel-frequency cepstral coefficients derived using the zero-time windowing spectrum for classification of phonation types in singing.

open access: yesJournal of the Acoustical Society of America, 2019
Existing studies in classification of phonation types in singing use voice source features and Mel-frequency cepstral coefficients (MFCCs) showing poor performance due to high pitch in singing.

semanticscholar   +1 more source

Newborns' Language Discrimination May Not Reflect Sensitivity to Speech Rhythm: Evidence From Computational Modeling

open access: yesDevelopmental Science, Volume 29, Issue 4, July 2026.
ABSTRACT Human newborns are able to discriminate between certain languages but not others. This ability has long been attributed to sensitivity to rhythm—the temporal regularities in speech of different languages. Here, we demonstrate through a series of computational simulations that this discrimination behavior can be achieved using no temporal ...
Ruolan Leslie Famularo   +3 more
wiley   +1 more source

On the Optimal Selection of Mel‐Frequency Cepstral Coefficients for Voice Deepfake Detection

open access: yesExpert Systems
ABSTRACT The continuous evolution of techniques for generating manipulated audio, known as voice deepfakes, and the widespread availability of tools that produce convincing forgeries have created an urgent need for reliable detection methods.
Sergio A. Falcón-López   +3 more
openaire   +2 more sources

Inter‐Model Feature Fusion for Robust Low‐Resource Speech Recognition

open access: yesApplied AI Letters, Volume 7, Issue 2, June 2026.
Our Self‐Supervised Feature Fusion (SSF‐FT) method enhances low‐resource speech recognition by adaptively combining features from self‐supervised models trained with Contrastive, Predictive, and Reconstruction objectives. This attention‐weighted ensemble delivers robust performance, particularly in acoustically challenging conditions, extending current
Ussen Kimanuka   +2 more
wiley   +1 more source

Feature Extracting in the Presence of Environmental Noise, using Subband Adaptive Filtering [PDF]

open access: yes, 2009
In this work, a new feature extracting method in noisy environments is proposed. The approach is based on subband decomposition of speech signals followed by adaptive filtering in the noisiest subbbands of speech.
Samad, Salina Abdul
core  

Mel-frequency cepstral coefficients of voice source waveforms for classification of phonation types in speech

open access: yes, 2019
Voice source characteristics in different phonation types vary due to the tension of laryngeal muscles along with the respiratory effort. This study investigates the use of mel-frequency cepstral coefficients (MFCCs) derived from voice source waveforms ...
Kadiri, Sudarsana Reddy   +3 more
core   +1 more source

Automatic Speaker Recognition Based on Mel-Frequency Cepstral Coefficients and Gaussian Mixture Models [PDF]

open access: yesMehran University Research Journal of Engineering and Technology, 2013
This paper investigates the task of SR (Speaker Recognition) for the state-of-the-art techniques. The paper initially presents the technical description of automatic SR, followed by the comparative analysis of a number of methods available for feature ...
Sheeraz Memon   +2 more
doaj  

Application of Music Data Visualization Technology in Music Appreciation Teaching

open access: yesEngineering Reports, Volume 8, Issue 6, June 2026.
The simulation environment is used to simulate real‐world music appreciation scenarios. DL is employed to preprocess music data, extract features, and identify rhythm information, which is then associated with visual design parameters to construct a parametric model.
Xiaowei Chen
wiley   +1 more source

On Compensating the Mel-Frequency Cepstral Coefficients

open access: yes, 2008
This paper describes a novel noise-robust automatic speech recognition (ASR) front-end that employs a combination of Mel-filterbank output compensation and cumulative distribution mapping of cepstral coefficients with truncated Gaussian distribution ...
For Noisy Speech, Eric H. C. Choi
core  

Home - About - Disclaimer - Privacy