Results 81 to 90 of about 2,981,759 (215)
Abstract Distributed wireless acoustic sensors, such as accelerometers, in smart water networks are increasingly used to detect leaks before they escalate into pipe breaks. In addition to leak‐induced signals, these sensors also capture acoustic noise generated by pumping and valve operations, pedestrians, traffic, and other environmental sources ...
Wei Zeng +5 more
wiley +1 more source
ABSTRACT Human newborns are able to discriminate between certain languages but not others. This ability has long been attributed to sensitivity to rhythm—the temporal regularities in speech of different languages. Here, we demonstrate through a series of computational simulations that this discrimination behavior can be achieved using no temporal ...
Ruolan Leslie Famularo +3 more
wiley +1 more source
Inter‐Model Feature Fusion for Robust Low‐Resource Speech Recognition
Our Self‐Supervised Feature Fusion (SSF‐FT) method enhances low‐resource speech recognition by adaptively combining features from self‐supervised models trained with Contrastive, Predictive, and Reconstruction objectives. This attention‐weighted ensemble delivers robust performance, particularly in acoustically challenging conditions, extending current
Ussen Kimanuka +2 more
wiley +1 more source
We propose a method for minimum mean-square error (MMSE) estimation of mel-frequency cepstral features for noise robust automatic speech recognition (ASR).
Tan, Zheng-Hua; id_orcid +3 more
core +1 more source
Fusion of Linear and Mel Frequency Cepstral Coefficients for Automatic Classification of Reptiles
Bioacoustic research of reptile calls and vocalizations has been limited due to the general consideration that they are voiceless.
Juan J. Noda +2 more
doaj +1 more source
On the Optimal Selection of Mel‐Frequency Cepstral Coefficients for Voice Deepfake Detection
ABSTRACT The continuous evolution of techniques for generating manipulated audio, known as voice deepfakes, and the widespread availability of tools that produce convincing forgeries have created an urgent need for reliable detection methods.
Sergio A. Falcón-López +3 more
openaire +2 more sources
Application of Music Data Visualization Technology in Music Appreciation Teaching
The simulation environment is used to simulate real‐world music appreciation scenarios. DL is employed to preprocess music data, extract features, and identify rhythm information, which is then associated with visual design parameters to construct a parametric model.
Xiaowei Chen
wiley +1 more source
Voice source characteristics in different phonation types vary due to the tension of laryngeal muscles along with the respiratory effort. This study investigates the use of mel-frequency cepstral coefficients (MFCCs) derived from voice source waveforms ...
Kadiri, Sudarsana Reddy +3 more
core +1 more source
Automatic Speaker Recognition Based on Mel-Frequency Cepstral Coefficients and Gaussian Mixture Models [PDF]
This paper investigates the task of SR (Speaker Recognition) for the state-of-the-art techniques. The paper initially presents the technical description of automatic SR, followed by the comparative analysis of a number of methods available for feature ...
Sheeraz Memon +2 more
doaj
Construction and Application of GAN Enhanced Virtual Interpretation Model for Sports Communication
This model is an end‐to‐end framework. Firstly, the style‐based generation network Style‐based GAN2 (StyleGAN2) is used to generate a highly realistic and adjustable static narrator portrait. Then, Bi‐directional Long Short‐Term Memory (Bi‐LSTM) is used to encode the Mel‐frequency Cepstral Coefficients (MFCCs), phonemes, and prosodic features of the ...
Li Zhang +3 more
wiley +1 more source

