Results 81 to 90 of about 3,612,587 (208)
An Underwater Acoustic Target Recognition Method Based on Restricted Boltzmann Machine
This article focuses on an underwater acoustic target recognition method based on target radiated noise. The difficulty of underwater acoustic target recognition is mainly the extraction of effective classification features and pattern classification ...
Xinwei Luo, Yulin Feng
doaj +1 more source
We propose a method for minimum mean-square error (MMSE) estimation of mel-frequency cepstral features for noise robust automatic speech recognition (ASR).
Tan, Zheng-Hua; id_orcid +3 more
core +1 more source
Flow chart of the base sound trace extraction of speech. ABSTRACT With the acceleration of internationalization, the deficiency of traditional English classroom in English spoken teaching is becoming more obvious, especially the lack of effectiveness and immediate feedback. Therefore, a spoken English assisted training model is proposed.
Lubing Shang, Lu Ma
wiley +1 more source
A robust deep learning model for fall action detection using healthcare wearable sensors [PDF]
The proposed technique begins with Butterworth’s sixth-order filtering of the data followed by segmentation through Hamming window application. The identification of essential patterns is achieved through the utilization of feature extraction methods ...
Abdulwahab Alazeb +6 more
doaj +2 more sources
Neural architectures for gender detection and speaker identification
In this paper, we investigate two neural architecture for gender detection and speaker identification tasks by utilizing Mel-frequency cepstral coefficients (MFCC) features which do not cover the voice related characteristics.
Orken Mamyrbayev +3 more
doaj +1 more source
Perkembangan kecerdasan buatan saat ini sangatlah pesat guna menciptakan sistem komputer yang mendekati perilaku manusia. Salah satu perilaku manusia yang dapat diadaptasi menggunakan kecerdasan buatan adalah kemampuan pendengaran manusia. Manusia mampu mengidentifikasi berbagai jenis sumber suara berdasarkan warna suaranya, termasuk suara alat musik ...
Sinta +2 more
openaire +2 more sources
On the Optimal Selection of Mel‐Frequency Cepstral Coefficients for Voice Deepfake Detection
ABSTRACT The continuous evolution of techniques for generating manipulated audio, known as voice deepfakes, and the widespread availability of tools that produce convincing forgeries have created an urgent need for reliable detection methods. This work considers the dimensionality of Mel‐Frequency Cepstral Coefficients (MFCCs) as a core design variable
Sergio A. Falcón‐López +3 more
wiley +1 more source
IMPROVED LONG-RANGE DEPENDENCY CAPTURE IN SPEECH SENTIMENT RECOGNITION: A COMPARATIVE STUDY OF MFCC-BILSTM AND MFCC-LSTM [PDF]
The accurate analysis of sentiments conveyed through spoken language is of paramount importance in the domains of human-computer interaction. Speech sentiment recognition requires capturing long-range temporal dependencies within the speech signal.
Suman Lata +3 more
doaj +1 more source
A playback speech detection algorithm based on log inverse Mel-frequency spectral coefficient
The popularity and portability of high-fidelity audio recording equipment and playback equipment poses a serious challenge for speaker recognition systems against playback attacks.Based on the differences between the original speech and the playback ...
Lang LIN +3 more
doaj +2 more sources
AUTOMATIC SEGMENTATION OF BROADCAST AUDIO SIGNALS USING AUTO ASSOCIATIVE NEURAL NETWORKS [PDF]
In this paper, we describe automatic segmentation methods for audio broadcast data. Today, digital audio applications are part of our everyday lives. Since there are more and more digital audio databases in place these days, the importance of effective ...
P. Dhanalakshmi, S. Palanivel, M. Arul
doaj

