Results 51 to 60 of about 3,813 (182)
This paper presents an approach for speech retrieval. The feature being used in this approach is MFCC. This approach does not use any phoneme recognizer or Speech to text tool hence it can be used for other languages as well leads to the problem of speech retrieval (SR).
Dr. Jyoti Srivastava +3 more
openaire +1 more source
Multi‐Modal AI Approach in Depression Detection and Treatment: A Systematic Review of Last Decade
Overview of multimodal approaches for depression detection and treatment. ABSTRACT Depression is a common and devastating mental health illness with serious personal and societal consequences. Despite advancing treatment techniques, there are still hurdles in the effective diagnosis and treatment of depression, such as prompt diagnosis, personalized ...
Smith K. Khare +3 more
wiley +1 more source
Gaussian Mixture Model Based Classification of Stuttering Dysfluencies
The classification of dysfluencies is one of the important steps in objective measurement of stuttering disorder. In this work, the focus is on investigating the applicability of automatic speaker recognition (ASR) method for stuttering dysfluency ...
Mahesha P., Vinod D.S.
doaj +1 more source
Seven‐stage AI‐powered voice biomarker pipeline for neurodegenerative disease diagnosis and clinical decision support. ABSTRACT Recent studies have shown that voice analysis provides a feasible, non‐invasive diagnostic alternative for the early detection of neurodegenerative diseases (NDs), such as Parkinson's disease (PD) and Amyotrophic Lateral ...
Abdulazeez Mousa +3 more
wiley +1 more source
This paper introduces two significant contributions: one is a new feature based on histograms of MFCC (Mel-Frequency Cepstral Coefficients) extracted from the audio files that can be used in emotion classification from speech signals, and the other – our
Muhammet Pakyurek +3 more
doaj +1 more source
Speech Preprocessing Method Based on Transmission Wave and Fluid Theory
With the continuous development of social economy, voice audio has become one of the important ways of streaming media, and is widely used in many fields. However, it is very important for how to process speech. In view of this requirement and limitation, this paper introduces transmission wave and fluid theory, tries to comb the corresponding noise ...
Lixian Wu +5 more
wiley +1 more source
This study aims to determine which model is more effective in detecting lies between models with Mel Frequency Cepstral Coefficient (MFCC) and Short Time Fourier Transform (STFT) processes using Convolutional Neural Network (CNN). MFCC and STFT processes
Dewi Kusumawati +3 more
doaj +1 more source
Speech and Language Markers of Bipolar Disorder: Challenges and Opportunities
ABSTRACT Background Clinicians aspire to predict the emergence of Bipolar Disorder (BD) in a timely manner. To accomplish this, markers reflecting mental states that can be gathered non‐invasively and at large scale are needed. Here, we systematically evaluate evidence relating speech‐based markers to mood states in BD.
Farida Zaher +4 more
wiley +1 more source
Scale-invariant MFCCs for speech/speaker recognition
The feature extraction process is a fundamental part of speech processing. Mel frequency cepstral coefficients (MFCCs) are the most commonly used feature types in the speech/speaker recognition literature. However, the MFCC framework may face numerical issues or dynamic range problems, which decreases their performance.
Tüfekci Z., Dişken G.
openaire +2 more sources
Weaving Intelligence: Thermally Drawn Multimaterial Fibers Toward AI‐Enabled Smart Textiles
Thermally drawn multimaterial fibers are rapidly advancing as intelligent structural units for next‐generation smart textiles. Integrating multimaterial architectures with neuromorphic and spiking‐neural‐network principles enables fabrics that can sense, compute, and adapt autonomously.
Vuong Dinh Trung +9 more
wiley +1 more source

