Results 61 to 70 of about 3,495 (182)

Fusion of Linear and Mel Frequency Cepstral Coefficients for Automatic Classification of Reptiles

open access: yesApplied Sciences, 2017
Bioacoustic research of reptile calls and vocalizations has been limited due to the general consideration that they are voiceless.
Juan J. Noda   +2 more
doaj   +1 more source

Application of Music Data Visualization Technology in Music Appreciation Teaching

open access: yesEngineering Reports, Volume 8, Issue 6, June 2026.
The simulation environment is used to simulate real‐world music appreciation scenarios. DL is employed to preprocess music data, extract features, and identify rhythm information, which is then associated with visual design parameters to construct a parametric model.
Xiaowei Chen
wiley   +1 more source

Automatic Speaker Recognition Based on Mel-Frequency Cepstral Coefficients and Gaussian Mixture Models [PDF]

open access: yesMehran University Research Journal of Engineering and Technology, 2013
This paper investigates the task of SR (Speaker Recognition) for the state-of-the-art techniques. The paper initially presents the technical description of automatic SR, followed by the comparative analysis of a number of methods available for feature ...
Sheeraz Memon   +2 more
doaj  

Construction and Application of GAN Enhanced Virtual Interpretation Model for Sports Communication

open access: yesEngineering Reports, Volume 8, Issue 6, June 2026.
This model is an end‐to‐end framework. Firstly, the style‐based generation network Style‐based GAN2 (StyleGAN2) is used to generate a highly realistic and adjustable static narrator portrait. Then, Bi‐directional Long Short‐Term Memory (Bi‐LSTM) is used to encode the Mel‐frequency Cepstral Coefficients (MFCCs), phonemes, and prosodic features of the ...
Li Zhang   +3 more
wiley   +1 more source

HSD-Net: a dual-branch CNN-BiLSTM network with hybrid cepstral fusion for heart sound classification

open access: yesFrontiers in Medical Technology
IntroductionCardiovascular diseases (CVDs) represent a major global health threat, making early detection crucial. While cardiac auscultation is cost-effective, its reliance on clinical expertise leads to significant variability in diagnostic accuracy ...
Caijian Hua   +3 more
doaj   +1 more source

Automated classification of vowel category and speaker type in the high-frequency spectrum

open access: yesAudiology Research, 2016
The high-frequency region of vowel signals (above the third formant or F3) has received little research attention. Recent evidence, however, has documented the perceptual utility of high-frequency information in the speech signal above the traditional ...
Jeremy J. Donai   +2 more
doaj   +1 more source

Optoelectronic Control of Redox Dynamics in POM Memristors for Noise‐Resilient Speech and Hardware‐Level Motion Recognition

open access: yesAdvanced Functional Materials, Volume 36, Issue 37, 7 May 2026.
Optoelectronic control of redox‐active polyoxometalate clusters in polymer matrices yields hybrid memristors with switchable volatile and non‐volatile modes, enabling reservoir‐type in‐sensor optical preprocessing and stable multilevel synapses for multimodal neuromorphic computing, including noise‐tolerant audiovisual keyword recognition and hardware ...
Xiangyu Ma   +13 more
wiley   +1 more source

A dual-stream CNN-based ConvMixer for durian ripeness classification using magnitude and phase features from knocking sounds

open access: yesResults in Engineering
Knocking-sound analysis provides a non-destructive method for assessing durian ripeness; however, most convolutional neural network methods mainly emphasize magnitude information while disregarding phase information that reflects the overall signal ...
Khomdet Phapatanaburi   +7 more
doaj   +1 more source

Comparison of cepstral coefficients to other voice evaluation parameters in patients with occupational dysphonia

open access: yesMedycyna Pracy, 2013
Background: Special consideration has recently been given to cepstral analysis with mel-frequency cepstral coefficients (MFCCs). The aim of this study was to assess the applicability of MFCCs in acoustic analysis for diagnosing occupational dysphonia in ...
Ewa Niebudek-Bogusz   +3 more
doaj   +1 more source

Multilingual Speaker Identification by Combining Evidence from LPR and Multitaper MFCC

open access: yesJournal of Intelligent Systems, 2013
In this work, the significance of combining the evidence from multitaper mel-frequency cepstral coefficients (MFCC), linear prediction residual (LPR), and linear prediction residual phase (LPRP) features for multilingual speaker identification with the ...
Nagaraja B.G., Jayanna H.S.
doaj   +1 more source

Home - About - Disclaimer - Privacy