Speech synthesis using Mel-Cepstral coefficient feature [PDF]
This thesis presents a method to improve quality of synthesized speech by reducing the vocoded effect. The synthesis model takes mel-cepstral coefficients and spectrum envelopes as features of the original speech waveform. Mel-cepstral coefficients could
Wang, Lu
core
The subject matter of the study is the analysis of the influence of pre-processing stages of the audio on the accuracy of speaker language regognition. The importance of audio pre-processing has grown significantly in recent years due to its key role in
Olesia Barkovska, Anton Havrashenko
doaj
Emosi merupakan perilaku manusia yang dapat diungkapkan dengan tingkah laku berupa raut wajah dan suara. Suara adalah suatu gelombang longitudinal yang merambat di udara. Pada kehidupan sehari-hari manusia berkomunikasi dengan menggunakan suara. Baik itu
Helmiyah, Siti +2 more
core
Emotion Recognition from Speech Signal Using Mel-Frequency Cepstral Coefficients
In this paper, mel-frequency cepstral coefficients are investigated for emotional content of speech signal. The features are extracted from spoken utterance.
ATASOY, AYTEN, KORKMAZ, Onur Erdem
core
A playback speech detection algorithm based on log inverse Mel-frequency spectral coefficient
The popularity and portability of high-fidelity audio recording equipment and playback equipment poses a serious challenge for speaker recognition systems against playback attacks.Based on the differences between the original speech and the playback ...
Lang LIN +3 more
doaj +2 more sources
A Comparative Investigation of Cepstral Feature Extraction Methods for Deepfake Speech Detection
The widespread adoption of voice-based authentication systems has been accompanied by an escalating threat from deep learning-based synthetic speech generation techniques.
Nida Akıncı, Erdal Özbay
doaj +1 more source
Klasifikasi jenis suara manusia menggunakan algoritma Convolutional Neural Network dengan metode ekstraksi Mel-Frequency Cepstral Coefficients [PDF]
Deteksi jenis suara manusia dalam konteks paduan suara sebagai penggunaan media pembelajaran dengan menggunakan CNN dan MFCC. Penelitian ini bertujuan untuk mengklasifikasi jenis suara manusia sehingga terbaginya ke 4 label yaitu sopran, alto, tenor, dan
Fatah, Muhammad Reza Abdul
core +1 more source
Evaluation of the Vulnerability of Speaker Verification to Synthetic Speech [PDF]
In this paper, we evaluate the vulnerability of a speaker verification (SV) system to synthetic speech. Although this problem was first examined over a decade ago, dramatic improvements in both SV and speech synthesis have renewed interest in this ...
Pucher, M. +4 more
core
Voice Disorder Classification Based on Multitaper Mel Frequency Cepstral Coefficients Features. [PDF]
Eskidere Ö, Gürhanlı A.
europepmc +1 more source
New time-frequency derived cepstral coefficients for automatic speech recognition [PDF]
The goal is to improve recognition rate by optimisation of Mel Frequency Cepstral Coefficients (MFCCs): modifications concern the time-frequency representation used to estimate these coefficients.
Wassner, Hubert, Chollet, Gérard
core +1 more source

